Fireworks AI Releases Fireworks Nexus: A Drop-In Routing and Cost-Control Layer That Moves Routine Coding Work to Open-Weight Models
Fireworks AI has released Fireworks Nexus, an AI management and routing platform aimed at engineering organizations. It connects the coding tools developers already use to a managed layer of open-weight models. The problem it targets is well documented. Forbes reported that Uber exhausted its entire...
A new field report shows how scientists use AI coding agents to modernize scientific computing, accelerating software development and discovery in genomics and beyond.
Fish Audio raises $50M seed to build AI voice models for creators and enterprises
Since launching last year, the startup today has more than 8 million people using the open-source or hosted version of its models, and now generates annual recurring revenue of $21 million.
Microsoft AI Releases MAI-Cyber-1-Flash: A 5B-Active-Parameter Cyber Model That Pushes MDASH to 95.95% on CyberGym
Microsoft AI has released MAI-Cyber-1-Flash, its first model built specifically for cyber defense. It is a 137B total, 5B active sparse MoE fine-tune of MAI-Code-1-Flash with a 256k context window. The model does not ship as a standalone endpoint — it runs inside MDASH, Microsoft's multi-model agent...
Deploying a 1-Bit Bonsai-27B Model with PrismML llama.cpp and OpenAI-Compatible Local Inference Workflows
In this tutorial, we deploy the 1-bit Bonsai-27B language model using the PrismML fork of llama.cpp, which provides the specialized CUDA kernels required to decode the model’s Q1_0_g128 GGUF quantization format
The post Deploying a 1-Bit Bonsai-27B Model with PrismML llama.cpp and OpenAI-Compatible ...
Kimi AI and kvcache-ai Open Sources ‘AgentENV’: A Distributed System that Powers Agentic Reinforcement Learning (RL) Training for Kimi K3
Moonshot AI's Kimi team and kvcache-ai open-sourced AgentENV (AENV) under MIT, as part of Kimi K3 Open Day. It runs agent sandboxes as Firecracker microVMs with millisecond snapshot, resume, and 16-way fork, behind an E2B-compatible API.
The post Kimi AI and kvcache-ai Open Sources ‘AgentENV’: A Dis...
Designing Skill-Driven Financial Analysis Agents with Claude, Python, MCP Connectors, and Automated Deliverables
In this tutorial, we build an advanced workflow around Anthropic’s financial-services repository and reproduce its skill-driven architecture in pure Python. We begin by installing the required libraries, cloning the repository, and programmatically mapping its agents, vertical plugins, partner integ...
OpenAI called the Hugging Face attack unprecedented. But we’ve been here before.
This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. Reading OpenAI’s account last week of how some of its models broke their containment and hacked into the computer systems of Hugging Face, another AI company, was...
OpenAI’s Hugging Face breach has reignited the debate over alignment and control
OpenAI's Hugging Face breach has reignited debate over AI alignment and control, exposing competing views on whether increasingly capable AI should be better aligned, better contained, or both.
Perplexity Releases pplx, a Single-Binary CLI That Puts Its Search API in the Terminal for Coding Agents
Perplexity has released pplx, an official command line client for its Search API. The tool exposes two commands — pplx search web and pplx content fetch — and returns exactly one JSON object on stdout. It ships as a checksum-verified single binary for macOS arm64 and Linux, alongside an Agent Skill ...
Google’s AI search is rapidly becoming the default, new data shows
Google’s AI Overviews now appear in 43% of searches, underscoring how quickly AI-generated answers are becoming the default way people discover information online.
Lightbits Labs Strengthens Enterprise Linux With Ubuntu Certification
Native Ubuntu Support Simplifies Deployment of High-Performance Software-Defined Block Storage for Private Clouds Built in Kubernetes and OpenStack Environments Lightbits Labs®, inventor of the NVMe® over TCP storage protocol and Inferra™, the first KV cache prefetch engine for AI acceleration, toda...
ARC Cuts Documentation Time by 18.5% and Optimizes Coding Accuracy Using Suki
One of Texas’s largest multispecialty groups achieves 97% clinician engagement rate — far exceeding industry benchmarks — as ambient clinical intelligence scales across 40 locations Austin Regional Clinic (ARC), one of the largest multispecialty medical groups in Central Texas, serving more than 700...
Coalesce Capital Announces Growth Investment in Workstreet
Coalesce Capital (“Coalesce”), a private equity firm focused on investing in next-generation technology-enabled services companies, today announced a strategic growth investment in Workstreet, (“the Company”) a leading provider of AI-native compliance and cybersecurity solutions to companies in regu...
92% of Healthcare Leaders Demand Clinical Expertise to Trust AI
New national survey finds adoption stalls for structural reasons, even as organizations see value Carta Healthcare, the leader in enterprise clinical data management, today released findings from a national survey of U.S. healthcare leaders showing that AI is proving its worth but failing to expand,...