Nous preprint demonstrates memory improvement certificates without source calibration
A preprint presents a method for certifying improvements in an agent's memory decisions without first estimating how reliable its information sources are. In one mathematical model family, learning and certifying useful decisions ...
Preprint reports 31% lower sea-ice prediction error by preserving expert uncertainty
A machine-learning preprint reports a 31% reduction in mean absolute error on sea-ice concentration compared with training on hard labels. Its method preserves the ranges supplied by individual experts instead of collapsing their ...
Preprint reports faster GB200 training with polynomial function replacements
Researchers report in an arXiv preprint that replacing selected mathematical functions with short polynomial calculations improved complete training-step throughput on GB200 hardware by 2.7%, 2.9%, and 8.0% across three tasks. Th...
Trillium Labs launches with plans to publish AI training experiments
Nathan Lambert and Tom Zick launched Trillium Labs, a nonprofit that plans to publish AI experiment details so outside researchers can study and replicate its work, WIRED reported. The lab has raised an undisclosed sum from Schmid...
Lean plans four proof-checking kernels to guard against AI exploits
Lean's next version will ship with four different proof-checking kernels instead of one, New Scientist reported. These components check the logic of mathematical proofs, and the change follows an AI-assisted stunt that exploited s...
Meta releases code to connect Muse with custom hardware
Meta released open source code that lets builders connect its Muse AI agent to their own hardware, The Verge reported. The SDKs support off-the-shelf ESP32 boards and Raspberry Pi devices, with connections to displays, buttons, se...
