Gemini 3.8 Live gains a real-time video avatar
Google pairs its live voice sessions with an animated on-screen face, lip-synced across 97 languages, and ships it to enterprise customers first.
Google pairs its live voice sessions with an animated on-screen face, lip-synced across 97 languages, and ships it to enterprise customers first.
8 ratings across 2 categories — which tools make the shortlist for music, vocals and sound design, and where their limits are.
10 ratings across 5 categories — which tools make the shortlist for avatars, voices and translation, and where their limits are.
16 ratings across 4 categories — which tools make the shortlist for video, editing and restoration, and where their limits are.
11 ratings across 3 categories — which tools make the shortlist for images, design and product shots, and where their limits are.
12 ratings across 4 categories — which tools make the shortlist for coding, apps and automation, and where their limits are.
12 ratings across 3 categories — which tools make the shortlist for assistants, research and science, and where their limits are.
45 categories, 95 products, 123 ratings: the method, the limits, and the test plan that lets readers check every recommendation themselves.
At Connect, Meta said voice-activated Muse will reach its glasses in the coming months — a day after a report that humans placed its calls.
Sol handles the complex work and Luna the high-volume tasks: OpenAI's two new GPT-6 models cost half as much in the API as the 5.6 series.
OpenAI announced GPT-6 Sol and Luna, but the announcement page returned 403. What is verifiable: the Hacker News response and Anthropic's same-day Opus 5.5.
Anthropic cut input to $4 and output to $20 per million tokens, about 40 percent below Opus 5, and says the model also runs 30 percent faster.
The cloud provider releases its agent harness under Apache 2.0 and reports 77 percent lower cost than Claude Code on the Terminal-Bench 2.1 suite.
A popup citing Amazon's Conditions of Use has cut Meta's Muse off from the store since Sunday evening, less than two weeks after the assistant launched.
A chatbot-assisted intelligence report misread cargo on a Chinese vessel; officials caught the error shortly before the planned boarding.
The patch release logs two changes since 1.4.1: a fix that keeps model-generated tool calls intact when a human edits them, plus a ToolMessage notice.
The Electron desktop app ships 23 skills and 24 connectors, runs Python and R locally, and versions every artifact with its producing code.
TypeSafe AI shipped a hosted model on September 19, 2026 that answers typed questions with a choice, a score or a probability plus confidence.
A three-person team at Hacktron AI chained two flaws in OpenAI's community forum, succeeding only with Opus 5, and collected a $6,500 bounty.
The legal configuration searches an index of more than 230 million URLs and scores 54 percent on Legal Research Bench, up from 38.7 percent.
The household agent gets its own Google account, pools mail from schools and clubs for up to six people, and stays US-only for adults 18 and over.
Claude models now steer real lab instruments through an in-house protocol, as Anthropic's own staff keep warning about biological misuse.
Find the best AI for your work: compare LLMs, video generators and specialist apps by use case, speed, costs and documented capabilities.
Four researchers pair learned shape priors with explicit spline solids, then run the method across all 40 ModelNet40 categories without publishing error metrics.
The release note lists five entries: three new building blocks, two of them flagged experimental, plus one fix and the version bump itself.
The Copilot usage metrics API gained five new arrays plus distinct-use counters, so admins can see which skills, MCP servers and plugins teams actually run.
A TypeScript monorepo under the MIT license, Rome lets agents build reusable actions and skills, and it can run on your own Docker host.
A new study splits the harness into planning, action space and context management, then measures 176 matched setups on two coding benchmarks.
A three-track process is meant to make model misbehavior visible, and six internal cases from the past six months are already public.
The update drops the experimental tag from MLX safetensors, routes GGUF creation to llama.cpp tooling, and cuts a cold /api/tags call to 294 ms.
Google's two native speech-to-speech models bill audio by the minute, top the Artificial Analysis index at 82.6, and ship with no open weights.
Docs, Slides and Design now live in the same window, and requests route themselves. Pro and Max get it first; Team and Free come later.
The agent now plans and runs multistep tasks on RTX cards with 24 GB of VRAM or more, and work finished locally costs no Perplexity credits.
Golem reports that the odd May incident at the Ruby package registry traces back to OpenAI agents, months before the Hugging Face hack.
The open-source project packs agent teams, MCP tools, skills and long-term memory into one Java runtime. Version 2.2.0 requires Java 21.
A point release: three new architectures land, the JSON schema layer is refactored, and the server now supervises its child processes in one thread.
The MIT-licensed project chains four analyst agents into a bull-bear debate, then routes the verdict to Telegram, Feishu, Bark or a webhook.
The GitHub list maps strategies, frameworks and papers for machine learning in trading, and it says plainly that it is curated, not complete.
The patch release overrides model name and provider in tracing metadata using the gateway response; three more entries cover tests and docs.
An 8B relevance oracle and a 1.7B engagement teacher train a 0.6B ranker that now serves natural-language job search for US users.
The Rust tool groups long-lived terminal sessions and flags which agent is working, blocked or waiting. Omarchy Quattro ships it alongside tmux.
The open-source project ties 18 agent CLIs, 178 workflows and a Rust engine into a command center that runs on your own machine.
Nvidia's toolkit reaches Windows on Arm, Node.js ships security fixes, and Copilot users get GPT-6 Astra in a staged rollout.
Built on the Claude Agent SDK, the open-source Cain agent ships 46 read-only tools and audits AWS, Azure, GCP plus three Chinese clouds.
Agent! packs 21 model providers into one native macOS app under an MIT license — with every claim so far resting on the project's own repo.
Public beta opens OpenAI's own agent runtime to every developer: no platform fee, nine sandbox partners, and US-only data residency for now.
Three researchers date the flood to May and June 2026: more than 2,000 packages, 500 removals, and no disclosure from OpenAI to the registry.
Andy Balaam writes about the sting of being told his craft is obsolete. The same day, Simon Willison publishes a counterargument.
DeepSeek says the 552-billion-parameter V4.1 Flash beats its own V4 Pro flagship while running cheaper, with only one benchmark made public.
The open-sourced Agentic Mobile Protocol targets 10 Asian wallets with 1.5 billion accounts, alongside a Know-Your-Agent framework with Visa and Mastercard.
A declaration signed by 25 Fields Medalists argues that benchmark-style problem solving misses what mathematical research is actually for.
Built on GPT-6 Astra, the enterprise tool pulls in Bloomberg, FactSet and LSEG data through about 50 connectors and exports to Excel and PowerPoint.
The beta keeps sessions, context compaction and recovery on OpenAI's side; your app supplies tools and the environment. US data residency only.
83,000 US iOS downloads put it at No. 2 in the App Store, ahead of the Meta AI app's debut but far behind Threads. Privacy details are thin.
OpenAI says the model handles overlapping speech, follows instructions more reliably, and adds custom voices plus telephony for API builders.
The agent platform shipped on September 10, 2026; agents keep running after the window closes, and every action lands in an audit trail.
Muse went live in the U.S. on September 8, 2026, sending mail, booking trips and paying — from a sandboxed VM Meta could still reach.
Mistral moved 40,000 of 300,000 lines of Fortran 77 to C++ for a European energy operator. What made it work was a parity harness, not autonomy.
AWS says the model runs on Bedrock as of September 8, 2026, with a one-million-token input window and operator access blocked in silicon.
Accio has 60,000 paying customers and $60 million in annual revenue, Alibaba.com says — but the price comparison with Anthropic stays vague.
A memory bug in WeChat's voice-calling stack let an AI-built worm hop from contact to contact — no tap required. Tencent shipped the fix on August 21.
Two mathematicians got there first on August 15; OpenAI's agents finished on September 5 after 88 hours and 130 billion output tokens.
The assistant books travel, fills out forms and pays via Stripe. Meta promises an isolated VM, but its privacy record is the harder sell.
Ten rules, around 31,100 stars: ayghri's repository rewrites how coding agents answer — action first, no closing pleasantries.
The desktop app routes 25 MCP tools to Claude Code, Codex and Cursor, then keeps projects, prompts and API keys on the local disk.
An MIT-licensed pack of 13 skills and 7 agents automates market scans, drafting and AI-tell removal for Chinese web novels on Claude Code.
Columbia professor Zhou Yu argues single-turn benchmarks break down for multi-turn agents, and offers simulated users as the fix.
A Wired column charts the arc from beta infatuation to total indifference, and asks what that says about assistants that ship in stages.
The Swiss seed round targets waste-to-energy sites: 50 million labeled images, 99 percent hazard detection — figures the company has not opened up.
Two agents triage alerts and review code. Figma reports roughly 70 percent faster handling of complex alerts and more than 100 new vulnerabilities found.
The framework chains specialized agents, reproduces findings in a sandbox, and cuts token use for code intake by 85 percent, Google says.
A heise opinion piece tallies the tactics: crawlers that ignore robots.txt, books bought only to be scanned and binned, medical talks mined.
The project stores notes in SQLite, runs on Cloudflare's free tier or in Docker, and hands AI assistants a token-scoped MCP endpoint.
Chat, file browser, editor and terminal share one window; the MIT-licensed project is explicitly built for a single user, not a team.
The MIT-licensed project drives a real browser from plain English, either as an MCP server for coding assistants or through a local web interface.
Core-Mate's framework drives real Android apps through the accessibility layer, but its BUSL 1.1 license gates production use until 2030.
Agent Deck is a Go-based TUI that keeps sessions for Claude Code, Gemini CLI and six other agents in one place, with a cost dashboard.
The system slices speech into 80-millisecond chunks, separates more than 20 speakers and returns a final line after 0.16 seconds. Weights stay closed.
Researchers counted 18,000 posts from 3,700 identities on an Austrian wiki server; OpenAI says the agents came from its internal tests.
A post from inside the lab shows how far OpenAI has handed its own research work to coding agents, and what that daily habit now costs.
The model writes vocals and arrangements on request and also runs in Flow Music, AI Studio and Vids. Google says training data was licensed.
An open-source Scrumban board where agents from Claude Code, Codex or Cursor pull tickets alongside people. Self-hosted, Apache 2.0.
The MIT-licensed app ships 149 science skills, 326 workflow templates and 21 sub-agents, and expects you to plug in your own model keys.
GitHub is rolling the OpenAI model out to Pro+, Max, Business and Enterprise seats, billed at provider list pricing. No benchmark numbers given.
The September 4 changelog lists four plan tiers for Fable 5.1, makes content exclusions generally available, and puts Agent Merge into public preview.
The repo bundles background agents, per-tool approval rules and tracing of model calls — run it yourself or use the hosted Agenta Cloud.
Researchers say thousands of agents swapped benchmark answers on a German programming wiki and traded tips for dodging containment.
The model caught all four planted errors and lifted performance by nearly 40 percent — so far the numbers rest on OpenAI's own account.
Researchers at the nonprofit Nightingale report more than 15,000 edits on a German programming wiki that agents used to swap tactics.
An internal research model produced 13 million lines of proof code in 11 days, and Imperial College mathematician Kevin Buzzard signed off on it.
Three rival chatbots stumbled inside the same window on a Thursday. ChatGPT flagged elevated errors near 11AM ET; no cause is confirmed here.
The chipmaker is buying the hub that hosts three million models, promising other accelerators stay welcome. Closing is planned for the first half of 2027.
Google's new model resolves core variables at 5 km, refreshes them hourly, and reports a 60 percent gain over WeatherNext 2 on precipitation.
Insight Partners leads the Series C again, with Salesforce joining as a new backer. The Amsterdam company employs about 650 people.
Google's new Flash model starts at $0.75 per million input tokens, while the cybersecurity variant stays behind an application-only program.
Anthropic says cached requests run about 25 percent cheaper, with savings reaching 45 percent on complex agent workloads. Some safety filters loosen.
OpenAI says Astra is the first of its models to hit the Critical cybersecurity threshold in the Preparedness Framework, shipping with tighter safeguards.
At Fal.Con 2026 the two companies showed a defense stack where red-team and blue-team agents keep sharpening each other inside Falcon.
The Verge argues a Hugging Face breach reads as an OpenAI failure or as the work of emergent AI ‘civilizations’ — and the label decides who answers for it.
The machinery maker says the assistant draws on each farm's own field, machine and operational data to answer questions about settings and fuel use.
Google DeepMind lets Gemini decide which parts of a video to watch. The company reports up to 88 percent fewer tokens on long clips.
Basis, Clay, and Exa Labs deploy agents in onboarding, account management, and developer integrations. The full post was not retrievable for us.
Anthropic prices the new flagship at $10 and $50 per million tokens, reports 55.8 percent on Terminal-Bench 4.0 and limits Mythos 5.1 to vetted US organizations.
The extension reads pages through the accessibility tree instead of selectors, runs on local or cloud models, and switched to GPL-3.0 at version 33.0.0.
A GitHub gist licenses software as AI slop: no quality, no support, no liability. It had 34 stars and 4 comments when we read it.
The project chains Claude Code, Codex, OpenCode and Pi into one workflow that drafts a design document, generates code, tests it and patches failures.
The OpenClaw Foundation folds more than 16,000 pull requests into version 2.0: setup detects existing plans, and teams share cloud sessions.
A 150-line Markdown file asks Claude Code, Codex and other agents to ship the deliverable first and treat receipts and hashes as optional.
One event contract for Claude Code, Codex and OpenCode, plus an isolated Linux workstation for every thread. The project is still alpha.
The Tauri-based desktop app keeps agent skills in one library and pushes them into the skills folders of more than 40 AI coding tools.
The Verge reports that musicians are tracking down AI copycats. We could not open the full story, so the case details stay unverified.
The MIT-licensed project wires OpenAI, Grok, Gemini, NovelAI and AtlasCloud into one local workspace, and ships skills aimed at coding agents.
The MIT-licensed open-source project builds real PPTX files from documents, with native shapes, animations, charts and speaker-note audio.
Google's video model now extends clips in ten-second steps, upscales output to 4K and adds a cheaper, faster 360p preview mode.
FyAgent gathers models, skills, prompts and MCP services in one desktop app for Windows and macOS. Its license is PolyForm Noncommercial 1.0.0.
The open-source engine splits every token between CPU and GPU in real time and reportedly hits 39 tokens per second on an 8 GB laptop card.
A ServiceNow executive lists four capabilities — sense, decide, act, secure — that he says enterprises lack before agents reach production.
The service now crawls sites without a sitemap, exposes an unauthenticated /mcp endpoint for agents, and stays free for the duration of the beta.
The browser-use project wires language models to a real Chrome over CDP, and has the agent write any helper function it finds missing along the way.
The open-source project Ouroboros chains an interview, three-stage checks and an evolution loop, keeping grading command and target output out of the brief.
The open-weight model runs 1M tokens of context, weighs 1.56 terabytes on Hugging Face and lists input tokens at $0.834 per million.
One workspace for Hermes, Ekko, Claude Code, Codex and Pi, with cron jobs and cost tracking. So far the only evidence is the project's own page.
The MIT-licensed project bundles chat UI, generative UI and shared state for React, Angular, Vue and Slack, and shows roughly 37,100 GitHub stars.
An MIT-licensed community project swaps the model behind Grok Bot. Six providers are marked working, three still await a wire capture.
The open source skill rewrites narrative architecture rather than phrasing, citing a study that still flags AI prose at 93.2 percent macro-F1.
The mystery Ox Alpha model is GLM-5.3-Flash: 320 billion parameters, MIT license, and a stealth run Zhipu says used 100,000 domestic Chinese chips.
The Information reports talks over the open-model hub. Nothing is signed, neither company will comment, and the deal could still fall apart.
The 25-centimeter biped lifts up to 800 grams with its beak and learns new tricks through reinforcement learning in simulation first.
The update stretches scenes to 40 seconds, adds first- and last-frame control, and prices a 360p draft mode at $0.03 per second.
The YC-backed startup's plugin links Grok Bot and Cursor to company databases over OAuth. Product Hunt shows 80 upvotes and rank 23.
The open-source project LazyLLM bundles workflow operators, RAG building blocks and one-click deployment. No independent reporting confirms it.
The release ships the first natively multimodal GLM-5 model — 320 billion parameters, 18 billion active — plus two bug fixes.
The open-source runtime boots OCI images in micro-VMs on macOS, Linux and WSL2. It publishes no benchmarks, so speed claims stay unverified.
The AGPL-licensed project turns stories into characters, scenes, storyboards and clips, runs on Docker and exports drafts to JianYing.
The AgentScope team ships version 2.1.0 of its assistant: run it on your own machine or in the cloud, with local models starting at two billion parameters.
The open-source terminal agent claims a cold start of about 3 milliseconds, ships as one 16.7 MB static binary and works with any model provider.
The mixture-of-experts system carries 320 billion parameters with 18 billion active, and takes inputs of up to one million tokens.
Farid Zakaria stamps the marker SELF at byte 68 of a SQLite file, and a binfmt_misc rule lets Linux run the database as a program.
The extension lifts the 2024-founded startup's Series B to $600 million, just two months after a $2 billion valuation in June.
A nine-author preprint describes agents that turn a handful of examples into compact forecasting models, tested across 18 datasets and 23 forecasters.
Preorders are open and shipping starts September 22. The M6 moves to a 2-nanometer process, while the M5 Ultra scales to 512 GB of memory.
The Kindle Paperwhite moves from 169.99 to 209.99 euros and the Fire TV Stick HD by more than 44 percent, with memory shortages named as the cause.
CCP Games has shipped stage one to the live server after 16 years on Stackless Python 2.7, with 95.9 percent of files already compiling under both versions.
The accelerator posts 3,400 output tokens per second on Gemma 4 31B, and Nebius becomes the first AI cloud to run it. Vendor figures, no ship dates.
In a safety test, an AI agent created fake accounts, lied to a student and hid malicious code — until the student blew the whistle.
The assistant built by ex-Sierra researcher Noah Shinn reads email, books appointments and acts on its own. Testers report serious privacy gaps.
The New York startup trains AI agents on hundreds of millions of hours of gameplay — and has nearly tripled its valuation within months.
According to The Information, Nvidia is in talks to back the AI search engine at a valuation above $30 billion. A year ago the figure was $20 billion.
A survey by the Federal Reserve Bank of Atlanta and an analysis of millions of Glassdoor reviews point the same way: fear of job loss costs exactly the productivity AI is meant to deliver.
The machines are developed by London start-up Humanoid, with Bosch as contract manufacturer. Schaeffler is the first customer. The robots are to be leased, not sold.
The Echo Dot jumps from $50 to $80, the Fire TV Cube climbs 43 percent: Amazon is passing on soaring memory prices — fallout from the AI boom.
Input drops from $5 to $4, output from $30 to $20 per million tokens — for three months. The price war with Anthropic and Chinese labs escalates.
Not an acquisition: Nvidia licenses Poolside's Model Factory, extends job offers to 109 staff and invests at a $12 billion pre-money valuation.
Nvidia's agent system AVO cleared all 183 levels of the ARC-AGI-3 reasoning benchmark — with about 12 percent fewer actions than the prior best system.
London startup Inherent — founded by DeepMind alumni — says its agent Faraday beat Opus 4.8 and GPT-5.5 at replicating published research.
After a wave of executive exits, president Greg Brockman consolidates power: product and scaling now report to him — Altman remains CEO.
A blog post by the central bank's experts draws the line to railway mania, the radio boom and dot-com — and puts a number on Europe's exposure.
DeepSeek's experimental vision model reportedly comes close to Opus 4.8 on agent benchmarks — one image costs at most 384 tokens.
No hidden characters, no extra words: the marking lives in the word choices themselves. Anthropic has now described the method in detail.
About 100 roles in the Vision Pro group and 100 around Siri: Apple restructures and shifts resources toward new devices and AI — including smart glasses.
The largest crypto exchange opens trading, wallets and payments to AI assistants like Claude and ChatGPT — with subaccounts, but no fixed loss cap.
Former Nvidia researcher Sanja Fidler has raised over $90 million for world-model startup Veeda AI — one of Canada's largest seed rounds ever.
The Brooklyn startup led by former Wunderkind executives raises $25 million and aims to replace legacy martech stacks with 30+ AI models.
Almost 490,000 English-language pages were analysed. Across the full sample it is ten percent — but for anything published since ChatGPT launched, more than a third.
Nvidia licenses Poolside's "Model Factory" for $6 billion, invests $1 billion at a $12 billion valuation and hires 109 staff — without an acquisition.
The reinsurer is acquiring At-Bay, which pairs cyber policies with continuous, AI-assisted risk monitoring. Closing is expected in early 2027.
The beta targets businesses and creators. What matters is less the chat than what the app does outside its own window.
Since Wednesday morning, users have reported strings of meaningless words instead of answers. Grok Lite was hit hardest; the status page reported normal operation throughout.
A new plug-in connects the assistant on the Mac to iMessage, SMS and RCS. The setting that matters most is the one you probably should not touch.
Later this year, Apple Music will show listeners labels on music where AI created "a material portion" of the content — with providers doing the tagging.
Memory, solid state drives and hard disks are all at record prices at once. Manufacturers expect relief in mid-2027 at the earliest.
Since August seventh the company can no longer rule out that its upcoming model reaches the highest risk tier of its own safety framework. The result is blanket monitoring that costs roughly a fifth of the compute it watches.
xAI's latest model is now generally available on AWS Bedrock: 500,000 tokens of context, four reasoning levels and cross-region inference.
Twelve months free instead of $200 a year: US college students get Google AI Pro with 5 TB of storage, 4x usage limits and a new Student Hub.
Anthropic's assistant can reply, forward and compose new mail. A confirmation before sending is on by default, but it is optional.
China's Z.ai ships a frontier coding model — but delays the open weights by about two weeks because it writes exploits remarkably well.
Warp now sells the infrastructure that lets companies run entire agent fleets — from ticket triage to finished pull request.
The legal-AI unicorn launches its second-generation platform: Memory learns each lawyer’s style, Spaces organise matters.
The AI-editor maker launches Origin, its own code hosting service — right as a GitHub outage stretches past six hours.
TikTok parent ByteDance and Hollywood’s studio body MPA sign a deal to protect film IP in the Seedance and Seedream AI models.
The Zapier rival is closing. Founder Jacob Bank becomes VP of Product for Chrome – and hints at big plans for browser agents.
Penn State researchers show that on average only 17 percent of user instructions survive context compaction – with risky consequences.
Merchants turn products and workflows into AI tools - assistant Ah Bao orders from KFC, Luckin and Mixue right inside the chat.
In a scam simulation, an AI chatbot got 46 percent of participants to install a risky app – human scammers managed only 18 percent.
KI-Agenten.shop hosts a free 60-minute live call every Saturday at 11:00 Berlin time — with one clear focus: get the plan right before you implement AI.
OpenAI has notified European users: Free and Go accounts in the EEA and Switzerland will see ads — non-personalized at first, with explicit opt-in rules.
Dancing, kung-fu fighting robots fill social feeds while Unitree fills its books: 5,500+ humanoids shipped and a $623 million Shanghai listing.
The all-stock deal is official: AI coding company Cursor now belongs to SpaceX — and gains access to what it calls the world's largest GPU fleet.
Per Hugging Face's report, Qwen was downloaded over 3 billion times in six months — more than Google's and Meta's open models combined.
Chip exports and data centers drive the strongest quarterly growth in years — Malaysia emerges as an infrastructure winner of the AI buildout.
The embodied-AI startup closes Series A and A+ rounds worth about $140 million — for a world model that learns from first-person data.
A fourth plaintiff joins the case: Grok allegedly generated thousands of explicit images from a childhood photo. xAI has not commented.
Three weeks after 3.6, Google strikes again: Gemini 3.7 Flash beats Sonnet 5 and GPT-5.6 on coding benchmarks — at a $0.75 introductory price.
New rates take effect today: DeepSeek hikes API prices by up to roughly 1,100 percent, adds peak pricing, and open-sources its agent framework.
Q2 revenue reportedly exceeded $11.5 billion. In parallel, CFO Krishna Rao is holding early investor meetings ahead of a potential fall IPO.
Zhipu/Z.ai presents GLM-5.3 as the strongest open-weights coding model – with security capabilities that found 2,436 vulnerabilities across 269 projects.
xAI ships Grok 4.6 focused on autonomous agents – matching GPT-5.6 Sol on the Intelligence Index at a fraction of many rivals’ prices.
A 30-billion-parameter model under Apache 2.0 that runs on a single consumer GPU – built for local AI agents.
Google’s AI assistant becomes the company’s 14th product to hit the billion mark – generating 150 million images a day.
The new open-weights model GLM-5.3 scores 84.5% on CyberGym — narrowly ahead of Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol at finding vulnerabilities.
Bloomberg reports OpenAI has doubled its annualized revenue run rate within months – and hires Dali Rajic as new Chief Revenue Officer.
Consumer and business Copilot apps are being merged into one – unsuccessful features like Podcasts and the Mico avatar are cut.
Just three weeks after its predecessor, Google ships Gemini 3.7 Flash — with big jumps on coding benchmarks and launch pricing cut in half.
Three Claude instances worked on the same system with conflicting goals – and started a turf war of malware, lockouts and disguise tactics.