AI1 min read
Viral AI agent Instinct raises $1B Series C at a $10B valuation
Instinct has raised a $1 billion Series C, saying 'we're just getting started.' The company is now valued at $10 billion.
From TechCrunch AI
Blog
Daily notes on new models, LLM releases, agent frameworks and AI research, written from the sources we follow and delivered as a newsletter every day.
Get the morning and evening issues
Every new post of the day, in one email. Confirmation required.
AI1 min read
Instinct has raised a $1 billion Series C, saying 'we're just getting started.' The company is now valued at $10 billion.
From TechCrunch AI
Agents1 min read
NarrateAI delivers production-ready LLM quality assurance on Amazon Bedrock. This post details five techniques—adaptive pipeline orchestration, cross-account multi-model failover, real-time streaming evaluation, composite evaluation, and...
From AWS machine learning blog
How this blog is made
POST /api/v1/route with execute: true.Photo: Yan Krukau
AI1 min read
ElevenLabs, a company that turns text into human-like speech, is now valued at 22 billion dollars. It has 600 million dollars in annual revenue.
From TechCrunch AI
Agents1 min read
LangSmith now displays agent session trajectories, making it easier to see what an agent did and find issues faster.
From LangChain blog
Agents1 min read
LangSmith now offers smithtune, a CLI that turns agent trajectories into custom fine-tuned models. It handles data, training, and evaluation.
From LangChain blog
AI3 min read
Learn what intermediate 200 means in routing, how it impacts performance, and how to evaluate it effectively within your system.
From growth-engine
AI1 min read
The seven-year-old startup has raised a $350 million Series E to fuel its data-as-a-service approach.
From TechCrunch AI
Agents1 min read
Skills let you encode domain-specific procedures as reusable, portable instructions for agents, but a fluent answer doesn't prove the agent picked the right skill or followed it. Learn how to measure skill selection and instruction follo...
From AWS machine learning blog
AI1 min read
Leaders in AI discuss plans to slow AI development, but details are unclear. Some see potential risks, others see benefits for safety and regulation.
From TechCrunch AI
Agents1 min read
Jev is a new model that evaluates agents by giving typed answers instead of generating text. It is faster and costs less than traditional methods.
From LangChain blog
AI1 min read
Dario Amodei plans to slow AI growth using independent safety evaluators. Industry leaders support or push back on this idea.
From TechCrunch AI
AI1 min read
Chinese startup Manus is raising $500 million at a $4 billion valuation. It resumed independent operations after breaking off its merger with Meta.
From TechCrunch AI
AI1 min read
Crusoe raised $3.9 billion in a Series F round. Its valuation is now $30.9 billion.
From TechCrunch AI
LLMs1 min read
Reinforcement learning causes agents to invoke tools based on superficial prompt cues rather than task necessity, with spurious invocation rates rising up to 39 percent in controlled environments.
From arXiv cs.CL
LLMs1 min read
Researchers propose REALM, a framework using retrieval feedback to reorganize long-term memory in LLM agents. It achieves 75.97% accuracy on LoCoMo and 65.11% on LongMemEval.
From arXiv cs.CL
LLMs1 min read
A new pre-tokenizer framework factors orthographic variations into reversible opcodes, reducing vocabulary requirements by up to 16% across six corpora.
From arXiv cs.CL
LLMs1 min read
Researchers released NepKANUN, a fine-tuned LLM integrated with RAG to assist with Nepali legal texts. The system achieved F1 scores of 0.82, 0.77, and 0.71 on simple, moderate, and complex tasks respectively.
From arXiv cs.CL
LLMs1 min read
A new study maps 22 LLMs across closed-source and open-source categories using self-reported personality traits projected into a six-dimensional archetype space.
From arXiv cs.CL
LLMs1 min read
Research shows ordinary typos rotate hidden state vectors by 43 to 56 degrees, causing prompt injection probes to drop true positive rates by up to 12 percentage points.
From arXiv cs.CL
Posts are drafted from public feeds by models OpenSmartRoute routes to - the same router, skill and metering customers use - and always link to the original source. Corrections: support.
Archive (93)