AI category

LLMs News

Latest news on large language models — GPT, Claude, Gemini, Llama, and more.

  • LLMs
  • 40 recent stories
  • The AI Wrap

Latest LLMs stories

  1. OpenAI acquires presentation startup NextSlide

    TechCrunch · 2026-08-08

    NextSlide says its team members are now working on ChatGPT.

  2. Claude Vs ChatGPT: How These AI Assistants Differ

    Engadget · 2026-08-08

    In a practical breakdown of how Claude and ChatGPT AI models differ, one tends to fall short when it comes to quality responses and overall user experience.

  3. Anthropic sets Claude Code to Auto Mode by default to protect developers from bad approvals

    The Decoder · 2026-08-08

    Starting August 14, Anthropic will make Auto Mode in Claude Code the default for Pro, Max, and Team plans. The company says it's safer.

  4. Responding to the next frontier of critical cyber capabilities

    OpenAI Blog · 2026-08-07

    OpenAI is sharing preliminary cybersecurity evaluations for Astra and the steps we’re taking to strengthen safeguards and security controls.

  5. OpenAI says it slowed Astra model development over security concerns

    TechCrunch · 2026-08-07

    OpenAI said this model, which is still in development, reached its "critical cybersecurity threshold," meaning it could independently identify and carry out cyberattacks against

  6. Presentation: Keeping ChatGPT Fast as AI Development Accelerates

    InfoQ · 2026-08-08

    Martin Spier explains how agentic workflows dramatically increase code change volume at OpenAI. He d

  7. Claude Code sessions can now talk to each other and share context across terminals

    The Decoder · 2026-08-08

    Claude Code now lets sessions talk to each other. On macOS and Linux, instances running in parallel can send messages, share insights, and check on each other's status.

  8. Fields Medalist who published a paper on AI-driven human extinction now works for OpenAI

    The Decoder · 2026-08-08

    Newly awarded Fields Medalist Jacob Tsimerman is leaving the University of Toronto to join OpenAI and work on AI safety.

  9. How to Disable Gemini in Gmail and Google Docs

    WIRED · 2026-08-08

    New AI toolbars and prompts are showing up in Google Docs and Gmail. If you don’t want Gemini’s help in writing documents and emails, here’s how to turn that stuff off.

  10. xAI's Imagine Image 2.0 lands just behind OpenAI's GPT-Image-2 in Arena benchmarks

    The Decoder · 2026-08-08

    xAI has released Imagine Image 2.0 as a new image generator for Grok. The model ranks second in the Arena benchmarks, just behind OpenAI's GPT-Image-2.

  11. AI agents use roughly 600 times more energy than a simple chat prompt

    The Decoder · 2026-08-08

    Climate scientist Zeke Hausfather tracked his Claude Code usage over eight weeks: 3.2 billion tokens and about 170 kWh of data center electricity.

  12. Mistral AI Releases Shieldstral 1.0 3B: An Open-Weights Policy-Adaptive Multimodal Safety Classifier Matching Models 7× Its Size

    MarkTechPost · 2026-08-08

    Mistral AI has released Shieldstral 1.0 3B, an open-weights, policy-adaptive multimodal safety classifier that frames content moderation as a single yes/no question instead of a

  13. Mistral Is in the Right Place at the Right Time

    WIRED · 2026-08-04

    Open-weight AI models are having a moment in the wake of recent turmoil at US tech giants. For French AI lab Mistral, that’s the the best thing that could have happened.

  14. Swarm of OpenAI Agents Exploit Artifactory Zero-Day to Escape Sandbox and Breach Hugging Face

    InfoQ · 2026-08-04

    Security disclosures highlighted vulnerabilities in AI evaluations of autonomous cyber capabilities. Notably, OpenAI’s models escaped sandbox isolation, breaching Hugging

  15. Metro Bank customer fights for £14,000 refund after AI-linked fraud

    The Guardian AI · 2026-08-04

    Lender was told money was being taken without authorisation, with cash used to buy credits for Claude chatbot A Metro Bank customer has told of his fight to get more than £14,000

  16. Apple is getting this wrong

    OpenAI Blog · 2026-08-03

    OpenAI addresses Apple’s baseless lawsuit, corrects claims about its employees, and shares messages documenting what happened.

  17. House Homeland Security Panel Calls Altman In Over OpenAI Breach

    Unite.AI · 2026-08-03

    The U.S. House of Representatives' cybersecurity committee has asked OpenAI CEO Sam Altman for a briefing on the company's rogue AI agent that attacked AI platform Hugging Face

  18. How to Secure AI Agents, MCP Servers, and LLM Apps in Production

    MarkTechPost · 2026-08-03

    AI agents, MCP servers, and LLM apps break the core AppSec assumption that applications do what their code says.

  19. OpenAI Hack Could Have Been 'Way Worse,' Hugging Face CEO Says

    Bloomberg · 2026-08-03

    Hugging Face CEO Clément Delangue says a hack by OpenAI could have been "way worse" if not for the company's defensive measures.

  20. 'Concentration of Power' One of Biggest Risks in AI, Says Hugging Face CEO

    Bloomberg · 2026-08-03

    Clement Delangue, CEO of AI company Hugging Face, sat down with Bloomberg's Ed Ludlow to discuss OpenAI models' hack on Hugging Face last month and the future of AI regulation.

  21. Alibaba's New Model, Amazon's AI Win and Apple's Next Chapter | Bloomberg Tech 8/3/2026

    Bloomberg · 2026-08-03

    Bloomberg’s Ed Ludlow breaks down Alibaba's latest Qwen model, its biggest ever, claiming performance on par with Anthropic.

  22. Influencers draw backlash for attending OpenAI’s first luxury trip

    TechCrunch · 2026-08-03

    OpenAI’s first-ever influencer brand trip is sparking online backlash as tensions over the use of AI continue.

  23. Gemini Spark now has Chrome web-browsing capabilities

    Engadget · 2026-08-03

    Google's AI assistant can "use your logged-in accounts and saved passwords to handle tedious web errands."

  24. Alibaba's new Qwen model is also taking your job, but this time it's great

    The Decoder · 2026-08-03

    Alibaba is marketing its new AI model Qwen 3.8 with a video that shows the AI working while a person enjoys their hobbies.

  25. How I’d Learn AI Engineering in 2026

    Towards AI · 2026-08-03

    You can open Codex, Claude Code, or Cursor today, describe an app in English, and get a convincing result in minutes. The agent can…

  26. Congress’ favorite AI tool? ChatGPT

    TechCrunch · 2026-08-03

    House spending records show OpenAI's ChatGPT dominates paid AI use on Capitol Hill, with congressional offices relying on the chatbot to draft memos, summarize legislation

  27. Prompt, Context, Loop: The Three Engineering Layers Every RAG System Is Built On

    Towards Data Science · 2026-08-03

    Enterprise Document Intelligence [Vol.1 #M2] - Every RAG system is built in three engineering layers stacked on one LLM call: prompt (the call itself)

  28. How to keep your conversations with ChatGPT, Gemini, Copilot or Claude as private as possible

    ZDNet · 2026-08-03

    Worried about your personal AI chats being exposed? Here's how to tighten your privacy across several major chatbots.

  29. How Google used AI agents to find and fix 1,072 Chrome security bugs - in 60 days

    ZDNet · 2026-08-03

    With 3.5 billion active users to protect, Google is relying on Gemini to find Chrome security bugs fast - and before attackers do.

  30. The Download: reward hacking explained, and suspected Iranian cyberattacks

    MIT Technology Review · 2026-08-03

    This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology.

  31. Unicorn, pelican, Middle-earth: OpenAI co-founder Karpathy is looking for the next AI vibe test

    The Decoder · 2026-08-03

    One paragraph of "Lord of the Rings" in, 5,500 lines of code out. Andrej Karpathy had Claude Opus 5 turn Tolkien's opening into a 3D browser scene.

  32. How Claude Help Me Build My $200k+ ML Resume

    Towards Data Science · 2026-08-03

    How use Claude to craft an outstanding resume that lands offers

  33. Two teams solved the same quantum crypto problem using GPT-5.6 just three hours apart

    The Decoder · 2026-08-03

    Two research teams independently solved the same open quantum cryptography problem using OpenAI's GPT-5.6 Sol Ultra, submitting their papers just three hours apart.

  34. China’s Alibaba takes another swipe at America’s AI supremacy

    The Verge · 2026-08-03

    Chinese tech giant Alibaba released what it says is its largest and "most capable AI model to date," claiming performance rivaling the best systems from US frontier labs Anthropic

  35. LWiAI Podcast #253 - Opus 5, Gemini 3.6, Kimi K3, Hugging Face Hack

    Last Week in AI · 2026-08-03

    Anthropic releases Opus 5 promising Fable 5-like capabilities, Google Releases Three New Gemini A.I. Models, and more!

  36. Apple Siri AI vs. ChatGPT: Which AI Assistant Is Better?

    Bloomberg · 2026-08-03

    When Apple Inc. releases Siri AI this fall in iOS 27, it will instantly become the world’s most widely distributed artificial intelligence chatbot.

  37. Here’s why AI agents lie and cheat to reach their goals

    MIT Technology Review · 2026-08-03

    MIT Technology Review Explains: Let our writers untangle the complex, messy world of technology to help you understand what’s coming next. You can read more from the series here.

  38. OpenAI's super PAC is funding AI-generated news site attacking industry critics

    Hacker News · 2026-08-03

    OpenAI's Super PAC has been revealed to be backing an AI-driven news outlet that publishes content critical of individuals and organizations opposing the company's policy

  39. The Cyber Alchemist's Ghosh on AI Cyber-security risks

    Bloomberg · 2026-08-03

    Ajoy Ghosh, Founder and Chief Information Security Officer at The Cyber Alchemist, discusses his perspective on cybersecurity measures and solutions companies should consider

  40. Alibaba’s Qwen3.8-Max AI Model Claims Benchmark Scores Rivaling Anthropic

    Bloomberg · 2026-08-03

    Alibaba Group Holding Ltd. released its biggest ever AI model, claiming performance on par with global leader Anthropic PBC in the latest Chinese breakthrough to challenge US