AI news story

A Startup Says It Cracked AI's Decade-Old Math Limit — Its LLM Read 12M Tokens for $8

A Miami startup says it ran a long-context job that costs about $2,600 on Anthropic’s top model for $8 on its own LLM, read 12 million…

  • LLMs
  • Source: Towards AI
  • Published: 2026-06-21
  • Signal score: 4
  • 28 sources

Editor's take

A startup claims to have developed a large language model capable of processing 12 million tokens at a significantly reduced cost, purportedly costing $8 for a task that would run into thousands of dollars on established models like Anthropic's Claude 2.

This development directly addresses a major bottleneck in LLM deployment: the prohibitive expense of processing extended contexts, crucial for applications like in-depth document analysis or complex code comprehension. If validated, it could democratize access to powerful long-context AI for a wider range of businesses and researchers, potentially shifting the economics of AI-powered services.

Future developments will hinge on independent verification of the model's accuracy and efficiency across diverse tasks, not just a single benchmark. The ability to scale this cost-effective long-context processing and its generalization capabilities beyond the reported 12 million tokens will be key indicators of its true impact.

Signal score: 4

This event was corroborated by 28 independent sources. The signal score weighs cross-source corroboration, recency, source weight and topic salience. How we rank stories.

More LLMs stories

  1. OpenAI acquires presentation startup NextSlide

    TechCrunch · 2026-08-08

    NextSlide says its team members are now working on ChatGPT.

  2. Claude Vs ChatGPT: How These AI Assistants Differ

    Engadget · 2026-08-08

    In a practical breakdown of how Claude and ChatGPT AI models differ, one tends to fall short when it comes to quality responses and overall user experience.

  3. Anthropic sets Claude Code to Auto Mode by default to protect developers from bad approvals

    The Decoder · 2026-08-08

    Starting August 14, Anthropic will make Auto Mode in Claude Code the default for Pro, Max, and Team plans. The company says it's safer.

  4. Responding to the next frontier of critical cyber capabilities

    OpenAI Blog · 2026-08-07

    OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.

  5. OpenAI says it slowed Astra model development over security concerns

    TechCrunch · 2026-08-07

    OpenAI said this model, which is still in development, reached its "critical cybersecurity threshold," meaning it could independently identify and carry out cyberattacks against

  6. Presentation: Keeping ChatGPT Fast as AI Development Accelerates

    InfoQ · 2026-08-08

    Martin Spier explains how agentic workflows dramatically increase code change volume at OpenAI. He d