Comparing INT4 and NVFP4 Palettes on Real Gradient Tensors
Four-bit training quantizes every number to one of 16 values. NVFP4's menu is {0, ±0.5, ±1, ±1.5, ±2, ±3, ±4, ±6} , with one scale factor per block…
Tech news from the best sources
Four-bit training quantizes every number to one of 16 values. NVFP4's menu is {0, ±0.5, ±1, ±1.5, ±2, ±3, ±4, ±6} , with one scale factor per block…
Originally published on Loop & Retry — field notes on building LLM agents that survive production. Most fine-tuning guides answer "how many exam…
Anthropic and OpenAI are racing to scale up while reducing dependence on Nvidia.
A new paper introduces a method to speed up reward-based fine-tuning by having the model generate a cheap, compressed copy of itself to draft text,…
A 35-billion-parameter model called Agents-A1 matches trillion-parameter models on multi-step agent tasks, according to a new paper from Shanghai AI…
Fine-tuning tests show "bias ... toward confidently representing the claims as true."
But training on "synthetic stories" that model good AI behavior can help.
Overtuning can cause models to "prioritize user satisfaction over truthfulness.”