Skip to content

Large Language Model

Topic archive5 matches

Back to homeGEO summary endpoint

2026-09-12

Technology

  • Moonshot AI releases Kimi K2.8 model with million-token context window: Moonshot AI has released its Kimi K2.8 model, which delivers performance approaching the K3 model. The new model features a million-token context window that is now open to all users. The release comes as the company prepares for an IPO in Hong Kong.

    Moonshot AI量子位

    Permalink
  • LogiMed-RoB benchmark reveals error compounding in LLM medical logic: Researchers have introduced LogiMed-RoB, a benchmark based on Cochrane Risk of Bias 2.0 expert logic to evaluate large language models across 860 randomized controlled trials. Testing on 10 state-of-the-art models revealed a severe error compounding effect, despite the top model achieving 98.88% atomic consistency.

    AI Models and ApplicationsarXiv

    Permalink

2026-09-07

Technology

  • More capable LLMs can increase systemic risk in financial markets: A study shows that improving individual large language model capability can degrade system-level outcomes rather than improve them. Researchers hypothesize that shared training and architectures cause more capable LLMs to behave similarly, creating correlated actions that do not diversify away.

    ResearcharXiv

    Permalink

2026-09-05

Technology

  • Anthropic's Claude formalizes Fermat's Last Theorem proof in Lean: Anthropic announced that its Claude AI model worked largely autonomously over an 11-day period to formalize the proof of Fermat's Last Theorem. The project resulted in the first complete computer-checked proof of the famous mathematical theorem using the Lean programming language.

    AnthropicTechmeme

    Permalink