ieee p3109 arithmetic formats for machine learning
the ieee p3109 draft standard defines parameterized binary floating-point formats and operations to support efficient, consistent machine learning computations.
topic
the ieee p3109 draft standard defines parameterized binary floating-point formats and operations to support efficient, consistent machine learning computations.
a semiotic scaffolding called peel reveals systematic distortions in ai-generated text condensations, showing fluency does not equal fidelity.
a large-scale analysis across four eeg datasets examines how different scalp regions contribute to predicting cognitive workload, revealing consistent patterns and practical implications for sensor selection.
a guide to five foundational papers covering transformer architecture, few-shot learning, scaling laws, instruction tuning, and retrieval-augmented generation for understanding large language models.
a new benchmark uses public prediction-market and blockchain data to evaluate how well models predict individual beliefs and trades.
google research releases its ai flood forecasting framework on github, enabling national agencies to train and customize models with local data.
direct preference optimization reduced text degeneration by an average of 59.4% across five ocr model families by using the model's own failure outputs as rejection pairs.
a new framework lets human agents approve or reject algorithmic price suggestions, using old pricing data to skip the slow start typical in sparse booking markets.
a new method reduces the cubic complexity of gaussian processes with gradients by using exact gradient reduction and vecchia approximation.
a position paper argues that in high-dimensional settings, many different mechanisms can produce the same data, so predictive success does not prove a model has found the true mechanism, and large language models can hide this by giving a single fluent explanation.
aura-mem uses a learned gate to write only when observations change actions, keeping memory fixed at 4,224 bytes regardless of episode length.
periodic and soft target updates can guarantee convergence in linear q-learning under explicit spectral and step-size conditions.
a new method uses gradient tests instead of validation loss to decide when to stop training gradient boosted trees, avoiding the need for a patience parameter.
a new diagnostic reveals that common anomaly detection benchmarks become unreliable when held-out classes overlap with normal data in representation space.
a new framework aligns structured electronic health record representations with large language models to improve clinical prediction and reasoning.
microsoft announced two new language models, mai-thinking-1 and mai-code-1-flash, with low active parameter counts and claims of clean training data.
an overview of advances in making large language models more interpretable through dynamic evaluation, statistical methods, and accessible tools.
microsoft's open source assert framework uses ai to turn plain-language rules into scored tests for application-specific ai behavior.