50-Day Julia Programming Challenge by Amey Thakur and Mega Satish, focused on computational modeling, high-performance computing, and machine learning.
-
Updated
Feb 21, 2026 - Julia
50-Day Julia Programming Challenge by Amey Thakur and Mega Satish, focused on computational modeling, high-performance computing, and machine learning.
Neural ODE Transformer in Julia, trained via adjoint sensitivity methods. Matched-architecture vs. discrete Transformer on Penn Treebank: 119.9 val perplexity vs 113.7 discrete. Single-seed result; discrete wins at this scale as expected, continuous-depth advantage expected to emerge at larger scale.
Add a description, image, and links to the flux-jl topic page so that developers can more easily learn about it.
To associate your repository with the flux-jl topic, visit your repo's landing page and select "manage topics."