AI Assistants
50 items tagged with this topic
Recent
Older
Special treat for the @andy_pavlo fans out there: an in-depth discussion about what happens to…
Special treat for the @andy_pavlo fans out there: an in-depth discussion about what happens to databases when billions of AI agents hit them, his move to build a research lab at @ClickHouseDB, and plently of hot takes o…
When is an autonomous vehicle actually good enough? https://t.co/VQTbWP1Jfj https://t.co/MA8e3U…
When is an autonomous vehicle actually good enough? https://t.co/VQTbWP1Jfj https://t.co/MA8e3UAYgb
Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses
Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve the…
Together Link: open models in the harness you already use. Start with one command today.
Together Link brings frontier open models like GLM 5.3 and Kimi K3 into the coding agent your team already uses, cutting model spend by over 50%.
Import AI 475: Swarm scaling; Google DeepMind watermarks biology; and the AI science economy
Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If you’d like to support this, please subscribe. Subscribe now When should you use swarms? When you are in a hurry:…How…
[AINews] not much happened today
a quiet day.
Synthesis Superintelligence: from Semiconductors to Superconductors — Periodic Labs’ Liam Fedus and Ekin Dogus Cubuk
A special Science pod and Engineering pod crossover.. with Forward Deployed Engineering kicker!
[AINews] Claude Haiku 5.5 — better than GPT-6 Luna at the same pricing
yay small models
Can a Cloud-Native Harness Make Agents Reliable Beyond the Desktop?
Kubernetes co-creators Craig McLuckie and Joe Beda aim to bring agent harnesses fully into the cloud.
[AINews] Quasi-Riemann-Hypothesis: OpenAI publishes 722 math papers solving 90 of the top 500 open math problems; “the most significant moment” in >100 years of mathematics
Our head hurts.
Data is the Hard Part
Bharat Patel is currently Accenture’s AI and data lead for its defense portfolio, a former enlisted Navy sailor, and a veteran of the Army acquisition world, with experience at MITRE and across the Department of Defense.
It seems like anything you ship can now be decompiled and rebuilt by AI. I think it'll soon be…
It seems like anything you ship can now be decompiled and rebuilt by AI. I think it'll soon be incredibly hard to make money from software unless you have proprietary data, distribution, or some other edge. Traditional…
and because agents make it easier, almost everyone is operating outside their expertise at some…
and because agents make it easier, almost everyone is operating outside their expertise at some point but luckily, you can just ask the model to teach you what you don't know
Desktop AI apps are great but expose users to supply-chain attacks and catastrophic mistakes by…
Desktop AI apps are great but expose users to supply-chain attacks and catastrophic mistakes by agents. We’re building a powerful desktop experience with focus on security & reliability. Excited to partner with @Microso…
The compute needed for the stage of AI we’re about to enter is going to be insane. Personal age…
The compute needed for the stage of AI we’re about to enter is going to be insane. Personal agents, agent swarms defending enterprises, agents that review all of our code for security issues, agents that process nearly…
While everyone's talking about agents, I've been exploring how teams can use them to work bette…
While everyone's talking about agents, I've been exploring how teams can use them to work better together. Here’s my take, from OpenAI DevDay 2026. https://t.co/W3rVoiTo84
Why Every Traded Personal Agents for One Company Agent
Speaker 1 | 00:00 - 00:12 The hot thing right now is dots. They're persistent agents from OpenAI. I was there at Dev Day when they launched them. I interviewed Sam about them, and I've been using dots since before they came out. And now we…
[AINews] Reflection Beam - 501B-A23B American Open Model
A small win for US open source
[AINews] not much happened today
a quiet day.
what is your default/workhorse coding agent today, Oct 2026?
what is your default/workhorse coding agent today, Oct 2026?
working at a higher level of abstraction has always required understanding the lower level ones…
working at a higher level of abstraction has always required understanding the lower level ones, I don't think coding agents changes this
Cyber will be one of the most defining domains for AI in the coming years, and a huge area of f…
Cyber will be one of the most defining domains for AI in the coming years, and a huge area of focus for most enterprises. AI is going to now create a whole new level of work for security dealing with the increase of vib…
I hooked up our team claw to X to trigger work faster. Unassigned sessions are for anyone to gr…
I hooked up our team claw to X to trigger work faster. Unassigned sessions are for anyone to grab. Our agent looks who worked on the related code last and pings people on the server. Whole thing was a prompt and team se…
we couldn't have built the @every agent without Claude Managed Agents extremely fun to get to s…
we couldn't have built the @every agent without Claude Managed Agents extremely fun to get to sit down and chat about it: https://t.co/RAsiCZwlZd
The Every team runs as much work as possible through agents. To help the whole team share skill…
The Every team runs as much work as possible through agents. To help the whole team share skills when a new model comes out, they built a company agent on Claude Managed Agents that everyone uses in Slack. Once it caugh…
Advancing computer use with Ironclad
Learn how OpenAI and Ironclad are training and evaluating AI agents on complex contracting workflows to advance computer use for professional work.
[AINews] Pi 1.0, Pi Durable, and AIE NYC
the minimalist harness goes stable... and TypeScript!
The most common mistake teams make is to treat evals as an additional QA step once the agent is…
The most common mistake teams make is to treat evals as an additional QA step once the agent is built. AI products are fundamentally different. Your evals are your product spec.
Introducing gdp-ts: Ghosts of Departed Proofs for TypeScript. gdp-ts is a library, linter and A…
Introducing gdp-ts: Ghosts of Departed Proofs for TypeScript. gdp-ts is a library, linter and AI skill for safer API design. Under this contract, sensitive functions require 'proofs' that the caller performed an authori…
A decent amount of agent adoption and deployment is bottlenecked by being able to test, tune, a…
A decent amount of agent adoption and deployment is bottlenecked by being able to test, tune, and optimize agents on “real” work environments. The files agents need to access, the CRM environments, email, and more. It’s…
Harnesses from labs have an incentive to burn tokens, which means harnesses from startups have…
Harnesses from labs have an incentive to burn tokens, which means harnesses from startups have a real utility, like how Grep can automatically swap token burn into deterministic tested code that is repeatable just by ob…
It would be interesting to be able to run the Muse/Dot "computer" locally. Feels like that is a…
It would be interesting to be able to run the Muse/Dot "computer" locally. Feels like that is a better agent environment than increasingly locked down MacOS.
Academia is for Ambition — Alex Zhang, MIT
We catch up with RLM first author Alex Zhang, MIT PhD, on Jev, PhD masxing, and the future of harnesses.
Working on a new little project. The 𝚁𝙴𝙰𝙳𝙼𝙴 is fully written by hand, because it's for hu…
Working on a new little project. The 𝚁𝙴𝙰𝙳𝙼𝙴 is fully written by hand, because it's for human consumption. The documentation internals are AI English, because they're for agents. This little rule of thumb can make…
We’re already starting to see what kind of new jobs AI is creating. AI requires significant tec…
We’re already starting to see what kind of new jobs AI is creating. AI requires significant technical work and surrounding services to deploy into the economy. This means jobs for AI engineers that build applied AI prod…
dot has become my primary interface to AI over the last few weeks. my dot's name is boo. he's c…
dot has become my primary interface to AI over the last few weeks. my dot's name is boo. he's cute! in a year, i bet i'll still be interacting in this way with @ChatGPT—e.g. as an always-on persistent agent—but boo will…
One year in: How Microsoft Research Asia – Singapore is advancing research, partnership and talent for real-world impact
Since launching a year ago, the Microsoft Research Asia — Singapore lab has established a strong foundation, deepened collaboration across government, academia, and industry, and explored how frontier AI research can create real-world valu…
Was talking to my agents at the park then got scratched by a squirrel nature is telling me some…
Was talking to my agents at the park then got scratched by a squirrel nature is telling me something
AI agent adoption is still very bimodal right now. You have coding and coding adjacent tasks wh…
AI agent adoption is still very bimodal right now. You have coding and coding adjacent tasks which have taken off, and then everything else. And even within coding, you have a very wide continuum of adoption patterns, w…
How will the world change if coding agents get 10x better?
How will the world change if coding agents get 10x better?
It'll be interesting to see how this innovator's dilemma plays out in the next few years as per…
It'll be interesting to see how this innovator's dilemma plays out in the next few years as personal agents become inevitable.. Amazon is protecting it's "ads" cash cow and *cough* the brand. Doordash, albeit smaller, i…
It's time to get serious about Security x AI. One of the joys of doing AIE is helping others st…
It's time to get serious about Security x AI. One of the joys of doing AIE is helping others start their own high quality conferences in their countries and domains. Every single 2025 partner has come back — and the 2nd…
Capy is like a 24/7 standup for your agents https://t.co/cNQwkWMxSz
Capy is like a 24/7 standup for your agents https://t.co/cNQwkWMxSz
My goal with GBrain: I want our agents to feel as native, expressive, and inevitable as the web…
My goal with GBrain: I want our agents to feel as native, expressive, and inevitable as the web eventually did—not as a chatbot feature, but as a personal Jiminy Cricket that knows you and helps you 24/7 always on, alwa…
If you are a home automation nerd, this is a fun prompt to try on Claude Code / Codex (edit bas…
If you are a home automation nerd, this is a fun prompt to try on Claude Code / Codex (edit based on your own setup): "Can you sniff network packets in our house and see what else can we automate? For context: We have a…
The future is verification-engineering. Proofs, (e2e) tests, benchmarks, linters… Some tests wi…
The future is verification-engineering. Proofs, (e2e) tests, benchmarks, linters… Some tests will be deterministic, some agentic. This looks great. https://t.co/ZSwn1UEupp
AI Podcasts be like: Cold open: "AI is about to destroy humanity. RSI is leading to a hard take…
AI Podcasts be like: Cold open: "AI is about to destroy humanity. RSI is leading to a hard take off that we won't be able to control." << but first, a message from our sponsors: meet AgentTina, your agentic workflow orc…