VentureBeat
-
Enterprise AI agents are only as reliable as the messiest documents behind them
VentureBeat
—
-
Enterprises winning with AI agents are limiting how much the agents can do alone
VentureBeat
—
-
Nvidia finds that simple linear math can replace costly AI model handoffs
VentureBeat
—
-
Slack wants to drag AI coding out of the terminal and into the group chat
VentureBeat
—
-
One in five enterprises can't stop a runaway AI agent's spending in real time
VentureBeat
—
-
NanoClaw comes to Slack, letting you create persistent AI agent teams and colleagues from a single message
VentureBeat
—
-
Serval’s super agent Catalyst creates roving background agents to identify and fix IT issues before they’re ticketed
VentureBeat
—
-
TrueFoundry's open source AI agent harness TrueForge boasts 30%-75% cheaper task completion than Claude Managed Agents
VentureBeat
—
-
VentureBeat names Rob Strechay as its first Lead Analyst, expanding its enterprise AI research push
VentureBeat
—
-
GLM-5.3 hits the API at $1.4/$4.4 per million tokens
VentureBeat
—
-
Block’s new Apache 2.0 agent workspace Berd works across models and harnesses, stores conversation history locally
VentureBeat
—
-
85% of companies burned by an AI mistake are racing to cut the humans who might catch the next one
VentureBeat
—
-
Commerce AI is fragmenting. Here is why that matters.
VentureBeat
—
-
Enterprises are overpaying for simple AI queries — Snowflake's gateway now auto-routes to cut costs up to 3x
VentureBeat
—
-
Qwen3.8-27B runs frontier-class coding agents and reasoning locally, no cloud API required
VentureBeat
—
-
Cursor launches Origin code hosting platform as GitHub outage exposes opening in AI coding race
VentureBeat
—
-
One AI module faked 86% of a pipeline's accuracy gains by feeding another the answers
VentureBeat
—
-
Enterprises with AI context layers report agent failures at more than twice the rate of those without one
VentureBeat
—
-
As enterprises confront AI agent sprawl, xpander wants them to own their own control and context layer
VentureBeat
—
-
How Heidi built production-ready AI for healthcare at global scale
VentureBeat
—
-
Cutting RAG inference costs 6x starts with deciding what never reaches the LLM
VentureBeat
—
-
DeepSeek's top-ranked V4 Flash stumbles on real agent tasks as its prices surge
VentureBeat
—
-
An eval harness found what qualitative review couldn't: AI models are most confident when wrong
VentureBeat
—
-
GLM-5.3 is here with advanced cyber capabilities — and reportedly already found a 'serious vulnerability' in Cursor
VentureBeat
—
-
Three Claude agents given conflicting orders sabotaged each other on a shared server — then didn't tell users what they'd done
VentureBeat
—
-
Google’s Gemini 3.7 Flash targets coding and agents with a 50% introductory price cut
VentureBeat
—
-
DeepSeek Harness launches as open source rival to Claude Code, alongside V4-Pro on API with higher prices
VentureBeat
—
-
Why Capital One built its multi-agent AI platform around open-weight models
VentureBeat
—
-
Writer says its new Palmyra X6 model cuts AI agent costs by 52% as token spending surges
VentureBeat
—
-
Four of five enterprises that secured AI agent identities still can't contain one that goes rogue
VentureBeat
—
-
SpaceXAI debuts Grok 4.6, overtaking Kimi K3's performance and matching GPT-5.6 Sol for world's third best on Artificial Analysis
VentureBeat
—
-
Skan AI raises $63 million betting that watching how employees actually work is the missing layer of enterprise AI
VentureBeat
—
-
Agentic security: Enterprises enforce agent permissions two-thirds of the time — and isolate high-risk agents less than one in five
VentureBeat
—
-
Agentic reliability and evaluations : Enterprises that got burned by a bad eval are the most likely to remove humans from the loop, not the least
VentureBeat
—
-
Agent context layers: Enterprises governing their AI data are catching twice as many bad answers as the ones who aren't
VentureBeat
—
-
Agentic orchestration: Enterprise AI organizations know how to govern agents but still can't meter what they cost
VentureBeat
—
-
Four AI agents coordinating in real time outperformed Claude Opus 4.8 on enterprise coding tasks
VentureBeat
—
-
Stanford is running 37,000 AI agents as a virtual biotech — and one of its drug designs got independently confirmed by Merck
VentureBeat
—
-
Tencent's Team Memory shares AI agent memory across a team — with no governance yet for when it's wrong
VentureBeat
—
-
No cloud, no GPUs, no problem: Liquid AI's new model LFM2.5-2.6B brings powerful AI agents to devices as small as a Raspberry Pi
VentureBeat
—
-
Qwen 3.8-Max and Claude Opus 5 show why raw benchmark scores don't predict the bill
VentureBeat
—
-
AI agents are part of your team now. Here’s how to secure all of them.
VentureBeat
—
-
The browser is where attacks land. Why is security still focused on the endpoint?
VentureBeat
—
-
Stop graphing everything: When GraphRAG actually beats vector RAG
VentureBeat
—
-
Structured AI data pipelines score 10.9 points below free-form code — DataFlow-Harness closes the gap
VentureBeat
—
-
How is your enterprise tracking AI agent telemetry? Groundcover thinks it should never leave your cloud
VentureBeat
—
-
Not just OpenAI: Now Anthropic says its internal models got online and cyberattacked 3 other organizations
VentureBeat
—
-
Thinking Machines debuts Inkling Small open source AI model nearing performance of predecessor at about 1/4 size
VentureBeat
—
-
AI price wars: OpenAI cuts GPT-5.6 Luna prices by 80% as model competition shifts toward cost
VentureBeat
—
-
Mastercard spent decades training its fraud system to see bots as thieves. Now bots are the ones doing the buying.
VentureBeat
—
-
Hush Security says the AI security problem has shifted from protecting models to governing identities as autonomous agents spread
VentureBeat
—
-
The lineage behind 69% of open models was never verified. Cisco just fingerprinted almost 900 for free
VentureBeat
—
-
NTT DATA AIVista and Snowflake: Identity alone won’t secure enterprise AI agents
VentureBeat
—
-
Companies are finally seeing AI ROI — and now they know how much more value it can deliver
VentureBeat
—
-
At Waymo, an AI project isn't ready until its evals are — not when the model performs well
VentureBeat
—
-
Enterprise AI agents can't talk to each other, can't be trusted with permissions, and can't be audited — 5 startups are already fixing that
VentureBeat
—
-
Nimble claims its new, domain-specialized Web Search Agents cut token costs in half while boosting retrieval accuracy
VentureBeat
—
-
Target SVP says its real AI moat isn't the models — it's everything built around them
VentureBeat
—
-
Bright Machines says its new hybrid robot cell could help solve a major AI infrastructure bottleneck
VentureBeat
—
-
Visa used Mythos to hunt for bugs in its own payment network, then open-sourced the harness that made it possible
VentureBeat
—
-
Instacart's CTO says AI made the company stop worrying about tech debt
VentureBeat
—
-
GM redesigned its engineering workflows around AI agents — and tripled its merged pull requests
VentureBeat
—
-
Runway couldn't fix a bug in its AI video model, so it turned the bug into a feature
VentureBeat
—
-
Snowflake launches Cortex AI Gateway to control AI agents and prevent runaway enterprise costs
VentureBeat
—
-
MCP just got its biggest update ever — here’s what changes for AI agents
VentureBeat
—
-
Fiduciary AI: Agents need to prove trustworthiness, not just ability
VentureBeat
—
-
Kimi K3's full weights are here, but they're 'open' with a caveat: What enterprises should know
VentureBeat
—
-
Microsoft launches AI cybersecurity model, agentic defense platform to cut enterprise security costs
VentureBeat
—
-
AI cites the deep pages but sends humans to the homepage — most sites are built backward
VentureBeat
—
-
Uh-oh: Some Claude shared conversations and Artifacts appear to be indexed and publicly accessible on Google Search
VentureBeat
—
-
Why SAP says enterprise AI agents need knowledge graphs and governance
VentureBeat
—
-
New ransomware targets AI model weights and can't even collect the ransom
VentureBeat
—
-
VentureBeat Research: Where enterprise AI agent governance hasn't caught up
VentureBeat
—
-
Anthropic launches Claude Opus 5, a cheaper AI model for coding, agents and enterprise workflows
VentureBeat
—
-
Microsoft launches new in-house AI models it says cut costs up to 89% versus OpenAI
VentureBeat
—
-
Agentic coding goes hands free as OpenAI brings GPT-Live's full duplex voice control to Codex and ChatGPT on the desktop
VentureBeat
—
-
Black Forest Labs launches FLUX 3 capable of generating images and 20-second video with audio — but in limited release to start
VentureBeat
—
-
Multi-turn attacks broke AI models 88% of the time — single-turn testing missed it, Cisco AI security lead warns at VB Transform 2026
VentureBeat
—
-
The AI compute gap: Enterprises are buying infrastructure faster than they can measure what it costs
VentureBeat
—
-
The agent security gap: 54% of enterprises have already had an AI agent incident, and most still let agents share credentials
VentureBeat
—
-
The AI context gap: Enterprise AI organizations have a trust problem, not a retrieval problem — and most are still building the fix
VentureBeat
—
-
The agent evaluation gap: Enterprise AI organizations have a reality-alignment problem, not a coverage problem — and most are shipping to production anyway
VentureBeat
—
-
Agentic orchestration: Enterprise AI organizations have a deployment problem, not a platform problem — and most are calling chatbots agents
VentureBeat
—
-
An AI now judges every move Rubrik's agents make, its AI chief said at VB Transform 2026 — but no one's measured if the judge is right
VentureBeat
—
-
The credential that let OpenAI's agents into Hugging Face exists in most enterprises right now
VentureBeat
—
-
AI agents aren't confidently wrong because of bad context — they're wrong because of bad data engineering
VentureBeat
—
-
OpenAI unveils Presence, a new platform that lets enterprises launch and manage realtime voice agents and chatbots
VentureBeat
—
-
Inflection AI returns to consumer market with Pi Journeys after Microsoft upheaval
VentureBeat
—
-
OpenAI's models broke containment and cyberattacked Hugging Face — what enterprises need to know
VentureBeat
—
-
Poolside drops Laguna S 2.1, an open-weight coding model that beats rivals 10x its size
VentureBeat
—
-
Stop adding more GPUs: Weka's new storage platform reduces load by caching 100% of an AI model's pre-calculated tokens
VentureBeat
—
-
Google's Gemini 3.6 Flash model cuts AI agent token costs by up to 65% on long horizon engineering tasks —and 3.5 Pro is on the way
VentureBeat
—
-
Google's Gemini Flash 5.6 model cuts AI agent token costs by up to 65% on long horizon engineering tasks —and 3.5 Pro is on the way
VentureBeat
—
-
Evals are the new PRD, Expedia’s AI chief tells VB Transform 2026
VentureBeat
—
-
Atlassian: Research shows organizations should approach AI at the team level, not the individual level, to achieve true ROI
VentureBeat
—
-
Atlassian: Why AI speeds up employees but not organizations
VentureBeat
—
-
Writer's AI harness cuts token spend nearly 40% — without sacrificing accuracy
VentureBeat
—
-
A single AI agent conversation can look perfect and still be broken, leaders from LangChain, Conviva and CoreWeave said at VB Transform 2026
VentureBeat
—
-
At VB Transform 2026, Zillow's engineering chief said AI ROI numbers only hold up if you measure before you build
VentureBeat
—
-
Safety guardrails blocked Hugging Face's defenders, not the attacker, when an AI agent breached its systems
VentureBeat
—
-
AI confidence just dropped 17 points in six months. That’s actually great news.
VentureBeat
—
-
The cleanup trap: Stop asking RAG to fix bad data
VentureBeat
—
-
Capital One releases VulnHunter, an open-source AI tool that finds software flaws before hackers do
VentureBeat
—
-
Intuit scrapped its own AI agent architecture twice in four months. At VB Transform 2026, its AI VP called that the fast path
VentureBeat
—
-
Agents think in milliseconds, legacy infrastructure doesn't. LinkedIn, Walmart and Zendesk shared how they closed the gap at VB Transform 2026
VentureBeat
—
-
Brex built its AI agent policy by watching what agents actually do, not by writing rules first
VentureBeat
—
-
China’s Moonshot AI releases Kimi K3, the largest open-source model ever, rivaling top U.S. systems
VentureBeat
—
-
The agent security gap: 54% of enterprises have already had an AI agent incident, and most still let agents share credentials
VentureBeat
—
-
The AI compute gap: Enterprises are buying infrastructure faster than they can measure what it costs
VentureBeat
—
-
Zero trust must now move at agent speed
VentureBeat
—
-
The AI context gap: Enterprise AI organizations have a trust problem, not a retrieval problem — and most are still building the fix
VentureBeat
—
-
The agent evaluation gap: Enterprise AI organizations have a reality-alignment problem, not a coverage problem — and most are shipping to production anyway
VentureBeat
—
-
Agentic orchestration: Enterprise AI organizations have a deployment problem, not a platform problem — and most are calling chatbots agents
VentureBeat
—
-
Thinking Machines open sources first multimodal language model, Inkling, focused on low cost and 'resistance to censorship'
VentureBeat
—
-
Amazon AGI director says AI agent reliability, not capability, is blocking enterprise deployment at VB Transform 2026
VentureBeat
—
-
Amazon AGI director says AI agent reliability, not capability, is blocking enterprise deployment at VB Transform 2026
VentureBeat
—
-
Cohere VP says enterprise AI sovereignty requires control of the full agent stack
VentureBeat
—
-
'We have maybe 20 months' to rebuild for AI agents, Meta's infrastructure VP tells VB Transform 2026
VentureBeat
—
-
Canva launches Code 2.0, offering AI website building to every user — including free accounts
VentureBeat
—
-
1Password moves into AI cost management, betting that token spend is the next enterprise budget crisis
VentureBeat
—
-
ACRouter picks the smartest AI model per task, beating Opus-only setups by 2.6x on cost
VentureBeat
—
-
The desktop infrastructure problem that kubernetes finally solves
VentureBeat
—
-
DeepSeek cut prices 75%. The 100x problem remains
VentureBeat
—
-
Forget typosquatting; slopsquatting is the software supply chain threat created by AI coding tools
VentureBeat
—
-
57% of enterprises have watched AI agents be confidently wrong. The fix is an agentic context layer, but who has one?
VentureBeat
—
-
57% of enterprises have watched AI agents be confidently wrong. The context layer is the reason why
VentureBeat
—
-
OpenAI introduces ChatGPT Work, a cloud-based AI agent that manages tasks across email, Slack and calendars
VentureBeat
—
-
Wall Street is debating the AI buildout. Enterprises just answered: 86% say their GPUs run at half capacity or less
VentureBeat
—
-
Enterprise AI is entering an evaluation gap: Agents are gaining autonomy faster than companies can verify them
VentureBeat
—
-
Google's TabFM skips per-dataset training and still predicts on tables it's never seen
VentureBeat
—
-
Shared API keys expose AI agents at 69% of enterprises, new VentureBeat research finds
VentureBeat
—
-
Enterprises using multiple AI models are underestimating failure rates by 2.25x
VentureBeat
—
-
The enterprise AI challenge nobody solves with code generation alone
VentureBeat
—
-
One interface isn't enough for enterprise AI
VentureBeat
—
-
SpaceX's Grok 4.5 launches at half the price of rivals — here's why that could rattle Anthropic and OpenAI
VentureBeat
—
-
OpenAI launches GPT-Live, a full-duplex voice upgrade that lets ChatGPT talk more like a person
VentureBeat
—
-
Slack’s Slackbot can now pull your CRM data, generate charts, and send DocuSigns — all from a chat message.
VentureBeat
—
-
AI has collapsed the cyber response window — resilience now starts before the attack
VentureBeat
—
-
The real cost, security, and culture problems behind enterprise AI agents
VentureBeat
—
-
The AI architecture that let Liberty Mutual shrug off the Fable 5 outage
VentureBeat
—
-
Box survey: Why enterprise AI leaders are outperforming their peers
VentureBeat
—
-
Anthropic brings Claude Cowork to mobile and web as usage data shows most users aren’t coding
VentureBeat
—
-
Digital-native startups are ditching rigid databases for their agentic stacks
VentureBeat
—
-
Anthropic's new "J-lens" reveals a silent workspace inside Claude that mirrors a leading theory of consciousness
VentureBeat
—
-
Tencent's Apache-licensed Hy3 takes on GLM-5.2 at half the size — and wins everywhere except coding
VentureBeat
—
-
What billions of AI predictions taught Expedia before the age of AI agents
VentureBeat
—
-
Build for the new AI era with Microsoft and NVIDIA
VentureBeat
—
-
How America's 250th birthday became a test of AI-powered collective intelligence
VentureBeat
—
-
Trunk Tools' stack cut document review from 60 days to 10 by ditching general-purpose models
VentureBeat
—
-
Enterprises lost Claude Fable 5 for a few weeks. New data shows two-thirds had already built their hedge
VentureBeat
—
-
New Alibaba AI framework skips loading every tool, cutting agent token use 99%
VentureBeat
—
-
Z.ai launches ZCode to challenge Cursor, Claude Code and GitHub Copilot in AI coding
VentureBeat
—
-
The Control Gap: Enterprise AI organizations have an ownership problem, not a technology problem — and most are governing it by hand
VentureBeat
—
-
Restaurants can now accept orders placed directly from ChatGPT and Claude thanks to Square's new, low-fee, no setup integration
VentureBeat
—
-
Anthropic is bringing back Claude Fable 5 globally after US lifts export control order — where can enterprises access it?
VentureBeat
—
-
Morgan Stanley cut its riskiest reconciliation job in half — by making its agents less autonomous
VentureBeat
—
-
Anthropic launches Claude Sonnet 5 at a steep discount to its top model as the company races toward a blockbuster IPO
VentureBeat
—
-
Google's Gemini Omni Flash hits the API, turning enterprise video production into a conversation
VentureBeat
—
-
Google unveils Nano Banana 2 Lite aka Gemini 3.1 Flash-Lite for low cost, 4-second fast enterprise image generations
VentureBeat
—
-
AI agents need context everywhere they run, even where the cloud can't follow
VentureBeat
—
-
Meituan open sources LongCat-2.0, the 1.6T, near-frontier agentic coding model that's been leading OpenRouter — trained entirely on Chinese chips
VentureBeat
—
-
Digital resilience compounds when AI and human expertise scale together
VentureBeat
—
-
DeepSeek open sources DSpark, a new framework to speed up LLM inference by up to 85%
VentureBeat
—
-
The attack that hijacked Claude Code came through Sentry. Datadog, PagerDuty, and Jira have the same exposure.
VentureBeat
—
-
Prompt injection is exploiting enterprise AI's biggest design flaws by targeting agents, RAG pipelines and model routers
VentureBeat
—
-
Claude Code turned every engineer into three. Now companies need more product thinkers
VentureBeat
—
-
New agentic memory framework uses 118K tokens per query. LangMem burns through 3.26M.
VentureBeat
—
-
Autonomous security agents need complete data. Here's how to check if yours is ready.
VentureBeat
—
-
OpenAI unveils GPT-5.6 Sol, Terra and Luna models — but only accessible to limited preview partners for now, per US Gov
VentureBeat
—
-
Most companies think they're building a software factory. They're actually just shipping bugs faster.
VentureBeat
—
-
Liquid AI's smallest model yet LFM2.5-230M beats models 4X its size at data extraction, can run 'anywhere'
VentureBeat
—
-
OpenAI's updated GPT-5.5 Instant is better at shopping, complex constraints, and understanding user intent — and it's already in the API
VentureBeat
—
-
Your enterprise AI agents should automatically remember which model is right for which task. Mindstone built the capability with Rebel
VentureBeat
—
-
Mistral launches OCR 4, turning document extraction into a full enterprise AI play
VentureBeat
—
-
Alibaba's model never trained as an agent — and improved agent performance across seven benchmarks
VentureBeat
—
-
Xiaomi's HarnessX rewrites its own AI scaffolding mid-task — and smaller models gain the most
VentureBeat
—
-
Stanford researchers will discuss their agentic 'scientists' that are on course to reshape drug discovery at VB Transform 2026
VentureBeat
—
-
How Shopify built an AI stack that doesn't care which models survive
VentureBeat
—
-
Amazon will present its framework for engineering trustworthy AI agents at VB Transform 2026
VentureBeat
—
-
Intuit will show off how it rebuilt its AI infrastructure to support fast and complex tasks at VB Transform 2026
VentureBeat
—
-
OpenAI unveils first custom AI inference chip, Jalapeño, with Broadcom — and its development was sped-up with OpenAI's own models
VentureBeat
—
-
Visa will offer an inside look at Project Glasswing and how the most powerful agentic models are changing enterprise security at VB Transform 2026
VentureBeat
—
-
Enterprise-grade AI image generation in 2 seconds is here: Krea 2 Raw and Turbo available as open weights under custom license
VentureBeat
—
-
Anthropic launches Claude Tag, replacing its Slack app with a persistent AI teammate that learns, monitors and works autonomously
VentureBeat
—
-
A proof of concept forgives a fragile data path. Operational AI does not.
VentureBeat
—
-
Alibaba's AI video model rises to No. 2 in global rankings, as OpenAI's Sora and ByteDance's Seedance fall away
VentureBeat
—
-
No Claude Fable 5? No problem: Sakana achieves frontier performance with new Fugu multi-model, auto synthesis system
VentureBeat
—
-
Why agentic enterprises need to become learning systems
VentureBeat
—
-
Researchers introduce Self-Harness, a framework that lets AI agents rewrite their own rules, boosting performance up to 60%
VentureBeat
—
-
AI hit the memory wall — now it needs a new context tier
VentureBeat
—
-
7,000 Langflow servers are under attack. LangGraph and LangChain have the same holes
VentureBeat
—
-
Fine-tuning forgets. RAG leaks context. Hypernetworks build the model your agent needs on demand.
VentureBeat
—
-
Anthropic's Claude Code Artifacts update brings live, shared dashboards and interactive workspaces to enterprises
VentureBeat
—
-
New AI optimization framework beats Claude Code and Codex by 2.5x on the same compute budget
VentureBeat
—
-
Copilot searched your mailbox. LiteLLM handed out admin keys. Run this 5-check audit before your stack is next
VentureBeat
—
-
Adobe embeds agentic AI workflows across Creative Cloud, shifting from media generation to production orchestration
VentureBeat
—
-
AWS enters the context layer race with a graph that learns from agents, not manual curation
VentureBeat
—
-
Anthropic ships major Claude Design overhaul with design system imports, code round-trips, and a fix for its token-burning problem
VentureBeat
—
-
Why Weibo’s tiny VibeThinker-3B has the AI world arguing over benchmarks again
VentureBeat
—
-
Z.ai’s open-weights GLM-5.2 beats GPT-5.5 on multiple long-horizon coding benchmarks for 1/6th the cost
VentureBeat
—
-
Databricks says it solved the decades-old data pipeline problem that's been slowing AI agents
VentureBeat
—
-
Stanford's DeLM cuts multi-agent task costs 50% — without a central orchestrator
VentureBeat
—
-
Satya Nadella warns that AI could hollow out entire industries, echoing the damage done by globalization
VentureBeat
—
-
When deep research isn't enough for your business: Sakana AI launches 'ultra deep research' agent for 100+ page reports in 8 hours
VentureBeat
—
-
85% of IT teams claim every AI agent is under control. Only 42% actually know who owns them.
VentureBeat
—
-
Vibe coding can build your pipeline. It can't explain it six months later
VentureBeat
—
-
Attackers scale deception with AI. Defenders need truth at machine speed.
VentureBeat
—
-
MCP solved tool calling. A2A solved coordination. What solves transport?
VentureBeat
—
-
Anthropic blocks all public access to Claude Fable 5, Mythos 5 following US government order — what enterprises should do
VentureBeat
—
-
Kimi K2.7-Code cuts thinking tokens 30% — but practitioners say the benchmarks don't check out
VentureBeat
—
-
Google researchers introduce 'faithful uncertainty', allowing LLMs to offer best guesses instead of hallucinations
VentureBeat
—
-
NanoClaw and JFrog launch 'immune system' to block AI agents from downloading malicious code
VentureBeat
—
-
PixelRAG beats text parsers on accuracy and cuts AI agent token costs 10x
VentureBeat
—
-
Microsoft’s open-source SkillOpt automatically upgrades AI agent skills without touching model weights
VentureBeat
—
-
Xiaomi's new open source, agentic AI coding harness MiMo Code beats Claude Code at ultra-long, 200+ step tasks
VentureBeat
—
-
Context compression finally works in production: new research cuts LLM input 16x without the accuracy hit
VentureBeat
—
-
What AI benchmarks miss about real-world performance
VentureBeat
—
-
Google's DiffusionGemma generates 256 tokens in parallel and self-corrects as it goes
VentureBeat
—
-
Why AI that works in the lab often fails in production — and what actually fixes it
VentureBeat
—
-
Surprise upset: GPT-5.5 beats Claude Fable 5 on brutal new Agents’ Last Exam benchmark
VentureBeat
—
-
Researchers say they trained a foundation model from scratch for about $1,500
VentureBeat
—
-
Anthropic CEO calls for FAA-style regulation of powerful AI models: what enterprises should know
VentureBeat
—
-
MassMutual's AI strategy: 12-month contracts, 30% productivity gains, zero lock-in
VentureBeat
—
-
Apple’s new Siri AI is more than just a smarter assistant — it's a new enterprise app layer
VentureBeat
—
-
Cohere open-sources a coding agent that runs on a single H100
VentureBeat
—
-
On-device AI agents hit a hard memory limit. Apple's new architecture routes around it.
VentureBeat
—
-
Anthropic brings Mythos to the masses with Claude Fable 5, its most powerful generally available model ever
VentureBeat
—
-
Every World Cup fan deserves a seat. Norton Neo says its free browser is the ticket
VentureBeat
—
-
AI is about to replace the interface. Business leaders aren’t ready
VentureBeat
—
-
Researchers trained an open source AI search agent, Harness-1, that outperforms GPT-5.4 on recalling relevant information
VentureBeat
—
-
Agentic AI solved coding — and exposed every other problem in software engineering
VentureBeat
—
-
When Claude changed, everything changed: Managing AI blast radius in production
VentureBeat
—
-
Microsoft AI chief says company was “set free” from OpenAI to pursue superintelligence
VentureBeat
—
-
Microsoft's AI Futurist explains how he uses Copilot — and the real-world problems enterprises are solving with agents
VentureBeat
—
-
AI agents are learning on the job — just not for your whole team
VentureBeat
—
-
Meta's AI support agent bound recovery emails for anyone who asked. Your SOC never saw an alert.
VentureBeat
—
-
Anthropic says 80% of its new production code is now authored by Claude — how your enterprise can keep up
VentureBeat
—
-
Google's new open source Gemma 4 12B analyzes audio, video — and runs entirely locally on a typical 16GB enterprise laptop
VentureBeat
—
-
Alibaba's Qwen3.7-Plus supports text, video and imagery inputs at low cost of $0.4/$1.6 per 1M token — but it's proprietary
VentureBeat
—
-
Perplexity AI unveils hybrid local-cloud inference system at Computex 2026
VentureBeat
—
-
The Agentic Reckoning: Enterprise AI organizations have a runtime problem, not a model problem — and most are building the wrong solution
VentureBeat
—
-
Enterprise AI agents keep creating data silos. Microsoft's Build answer is Microsoft IQ and Rayfin.
VentureBeat
—
-
Microsoft launches MXC, an OS-level sandbox for AI agents, with OpenAI and Nvidia already on board
VentureBeat
—
-
Microsoft debuts Surface RTX Spark Dev Box to run large AI models without cloud costs
VentureBeat
—
-
OpenAI's Codex update lets agents build interactive enterprise workspaces via Sites and role-specific plugins
VentureBeat
—
-
AI agents keep giving confident wrong answers. The context layer is enterprise AI's next production problem.
VentureBeat
—
-
Zip’s new AI agents want to stop your finance team from uploading contracts into personal ChatGPT accounts
VentureBeat
—
-
The design bottleneck for solo founders? AI has solved it.
VentureBeat
—
-
MiniMax-M3 debuts, eclipsing GPT-5.5 and Gemini 3.1 Pro on key benchmark performance for just 5-10% of the cost
VentureBeat
—
-
Anthropic’s browser agent got hijacked 31.5% of the time before safeguards engaged
VentureBeat
—