Large Language Model
Topic archive • 5 matches
2026-09-12
Technology
Moonshot AI releases Kimi K2.8 model with million-token context window: Moonshot AI has released its Kimi K2.8 model, which delivers performance approaching the K3 model. The new model features a million-token context window that is now open to all users. The release comes as the company prepares for an IPO in Hong Kong.
Moonshot AI • 量子位
PermalinkLogiMed-RoB benchmark reveals error compounding in LLM medical logic: Researchers have introduced LogiMed-RoB, a benchmark based on Cochrane Risk of Bias 2.0 expert logic to evaluate large language models across 860 randomized controlled trials. Testing on 10 state-of-the-art models revealed a severe error compounding effect, despite the top model achieving 98.88% atomic consistency.
AI Models and Applications • arXiv
Permalink
2026-09-07
Technology
More capable LLMs can increase systemic risk in financial markets: A study shows that improving individual large language model capability can degrade system-level outcomes rather than improve them. Researchers hypothesize that shared training and architectures cause more capable LLMs to behave similarly, creating correlated actions that do not diversify away.
Research • arXiv
Permalink
2026-09-05
Technology
Anthropic's Claude formalizes Fermat's Last Theorem proof in Lean: Anthropic announced that its Claude AI model worked largely autonomously over an 11-day period to formalize the proof of Fermat's Last Theorem. The project resulted in the first complete computer-checked proof of the famous mathematical theorem using the Lean programming language.
Anthropic • Techmeme
Permalink