Nous preprint demonstrates memory improvement certificates without source calibration
A preprint presents a method for certifying improvements in an agent's memory decisions without first estimating how reliable its information sources are. In one mathematical model family, learning and certifying useful decisions ...
Preprint reports 31% lower sea-ice prediction error by preserving expert uncertainty
A machine-learning preprint reports a 31% reduction in mean absolute error on sea-ice concentration compared with training on hard labels. Its method preserves the ranges supplied by individual experts instead of collapsing their ...
FP4 pretraining study reports faster throughput with training-loss tradeoffs
Researchers report in an arXiv preprint that a custom four-bit floating-point training route reached 37.9K tokens/s/GPU, compared with 18.8K for bfloat16 and 27.6K for Transformer Engine in matched tests on the same accelerator. T...
Trillium Labs launches with plans to publish AI training experiments
Nathan Lambert and Tom Zick launched Trillium Labs, a nonprofit that plans to publish AI experiment details so outside researchers can study and replicate its work, WIRED reported. The lab has raised an undisclosed sum from Schmid...
Lean plans four proof-checking kernels to guard against AI exploits
Lean's next version will ship with four different proof-checking kernels instead of one, New Scientist reported. These components check the logic of mathematical proofs, and the change follows an AI-assisted stunt that exploited s...
Anthropic commits $100 million to Claude engineer training academy
Anthropic launched Claude Frontier Academy on October 2 with a $100 million commitment and a goal of training 10,000 engineers by the end of 2027. Initial cohorts are running in San Francisco, New York and London, with participant...