Skip to content

Ai Model

Topic archive • 7 matches

Back to home • GEO summary endpoint

2026-09-21

Technology

  • Open Jev Models Are Here!!: A review of seven open-source Jev-style AI models examines their capabilities and performance. The evaluation utilizes tools and repositories including JevBench, SemIf, Bespoke-Nimble-9B, Decider, and Alex Wortega's OpenJev.

    Technology • Sam Witteveen

    Permalink
  • StepFun releases Step 5 Preview open-source model with 27B active parameters: StepFun has released Step 5 Preview, an open-source model featuring 27 billion active parameters. The model ranks among the top two open-source models in recent evaluations, demonstrating strong performance despite its compact active parameter size.

    AI Models • 量子位

    Permalink
  • Researchers introduce two-step test-time training protocol for VARC models: Researchers have introduced a novel two-step test-time training (TTT) protocol to improve rule induction in Vision ARC (VARC) models. The method first finetunes only the task embedding representing the transformation rule, and then freezes it to finetune the backbone.

    AI Research • arXiv

    Permalink
  • LogicTrack framework uses formal logic solvers to audit LLM reasoning steps: Researchers have proposed LogicTrack, a neuro-symbolic framework designed to verify the logical validity of intermediate reasoning steps in large language models. The system auto-formalizes each reasoning step into symbolic representations and verifies them using automated theorem provers.

    AI Research • arXiv

    Permalink
  • DeepSeek CEO calls training on Huawei chips one of its biggest bets: DeepSeek CEO Liang Wenfeng told investors that training its models on Huawei chips is one of the company's biggest bets. Huawei is reportedly scheduled to deliver its next-generation training chips to DeepSeek in the fourth quarter of 2026 or the first quarter of 2027.

    Artificial Intelligence • Techmeme

    Permalink
  • Unitree releases partial weights for 6-billion-parameter UnifoLM model: Unitree has released partial weights and data for UnifoLM-WLA-1.0, its 6-billion-parameter world-language-action model for humanoid robots. The model covers 64 tasks, though post-training code and full datasets are still planned for future release.

    humanoid robots • Pandaily

    Permalink
  • RBS-Attention training-free method optimizes long-context LLM prefill: To address the prefill bottleneck in long-context large language model inference, researchers proposed RBS-Attention, a training-free sparse-prefill method. The approach uses a centroid base branch to capture average relevance and a rescue branch to identify blocks at risk of underestimation due to mean dilution.

    AI Research • arXiv

    Permalink