Topics
Follow your own topics →
DIFF.BLOG
New Following Discover Jobs
More
Top Writers Suggest a blog Upvotes plugin
Report bug Contact About
Sign up
Menu
New Following Discover Jobs Top Writers
More
Suggest a blog Upvotes plugin Report bug Contact About
Sign up
The home for great developer writing.
We surface the best developer writing from thousands of independent blogs, updated daily.
Join now → Learn more
TOPICS

Accelerating Long-Context Model Training in JAX and XLA

188 · NVIDIA Corporation · Feb. 3, 2026, 5:43 p.m.
Agentic AI / Generative AI Developer Tools & Techniques Networking / Communications CUDA Graphs large language models jax XLA Machine Learning
Summary
This blog post discusses the advancements in training large language models (LLMs) that support extremely long context windows using JAX and XLA. It highlights techniques that help speed up the training process effectively, catering to the needs of developers and engineers working with LLMs.
Read full post on developer.nvidia.com →
MORE POSTS LIKE THIS
Writing an LLM from scratch, part 34b -- from bigrams to GPT-2, one component at a time (in JAX)
gpjt · Jul 8, 2026
large language models GPT-2
Optimizing for Low-Latency Communication in Inference Workloads with JAX and XLA
NVIDIA Corporation · Jul 18, 2025
Data Center / Cloud Development & Optimization
Controlling Reasoning Effort in LLMs
Sebastian Raschka · Jul 18, 2026
large language models Reasoning Modes
Overtraining as the path to human-like AI
seangoedecke.com RSS feed · Jul 18, 2026
AI large language models
Stochastic Nerds
Alecmuffett · Jul 7, 2026
uncategorised llm
Smarter data generation for faster Speculator training
Red Hat · Jul 6, 2026
large language models speculative decoding
Discover more posts →
AUTHOR
RECENT POSTS FROM THE AUTHOR
Choose how you want to continue.
Continue with GitHub Continue with Google