Research
50 items tagged with this topic
Recent
Introducing the Anthropic Cyber Mission \ Anthropic
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Import AI 475: Swarm scaling; Google DeepMind watermarks biology; and the AI science economy
Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If you’d like to support this, please subscribe. Subscribe now When should you use swarms? When you are in a hurry:…How…
Older
Synthesis Superintelligence: from Semiconductors to Superconductors — Periodic Labs’ Liam Fedus and Ekin Dogus Cubuk
A special Science pod and Engineering pod crossover.. with Forward Deployed Engineering kicker!
[AINews] Quasi-Riemann-Hypothesis: OpenAI publishes 722 math papers solving 90 of the top 500 open math problems; “the most significant moment” in >100 years of mathematics
Our head hurts.
China, the FCC, and the Logic of Transceivers
The dependencies that a potential rule on optical transceivers misses
Data is the Hard Part
Bharat Patel is currently Accenture’s AI and data lead for its defense portfolio, a former enlisted Navy sailor, and a veteran of the Army acquisition world, with experience at MITRE and across the Department of Defense.
China’s AI Safety Money Problem
China has plenty of AI safety talent, but weak philanthropy and a strong state shape what type of safety work actually gets done.
In Case You Missed It...
Best of September!
Moving @askalphaxiv to the dock is the best thing I have done this week.. Went from 0 to ~30 mi…
Moving @askalphaxiv to the dock is the best thing I have done this week.. Went from 0 to ~30 mins of reading papers on my phone whenever I’m waiting on things. https://t.co/aPwn0LgfcY
How Jump Trading is scaling quant research with ChatGPT
Jump Trading uses OpenAI to expand quantitative research. See how longer-running AI workflows combine multiple data sources with human review.
Sharing AI progress in mathematics
OpenAI publishes new results on open problems in mathematics from an internal frontier model and shares Lean proof formalizations and research details on GitHub.
Building China’s Two Starlinks
is two better than one?
Our approach to EU text provenance rules
How OpenAI is approaching text watermarking under EU rules. Learn where watermarks apply, how detection works, and why access starts with researchers.
Forecasting space weather risks on power grids
Extreme space-weather events can damage power systems on Earth and degrade GPS accuracy and satellite operations. A new machine learning system can predict where damage is likely to occur 30-60 minutes before a storm arrives. The post Fore…
China's Starlink Response
Crust it, and copy it
AGI Science Loops are coming, and there will be Muse/Instinct for those AGI Science Loops. Halm…
AGI Science Loops are coming, and there will be Muse/Instinct for those AGI Science Loops. Halmos is building that. https://t.co/By7Gr2jOj9
Academia is for Ambition — Alex Zhang, MIT
We catch up with RLM first author Alex Zhang, MIT PhD, on Jev, PhD masxing, and the future of harnesses.
Introducing Quine: An AI research system designed for the complexity of biology
Biology doesn't operate in silos, and neither should the AI representation of it. Quine is an early-stage research effort to create a multimodal world model of biology. By connecting insights across biological scales and modalities, Quine…
One year in: How Microsoft Research Asia – Singapore is advancing research, partnership and talent for real-world impact
Since launching a year ago, the Microsoft Research Asia — Singapore lab has established a strong foundation, deepened collaboration across government, academia, and industry, and explored how frontier AI research can create real-world valu…
Import AI 474: Platonic mindspace; TPUs in space; Zhipu starts an outer RSI loop
Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If you’d like to support this, please subscribe. Subscribe now Are minds patterns from a Platonic space, with bodies and…
I was paying almost $300/year for a YouTube research tool that shall go unnamed. The product wa…
I was paying almost $300/year for a YouTube research tool that shall go unnamed. The product was getting too complicated so I decided to check if Claude can build the core feature set that I actually needed. Claude buil…
A pop-up city where each building is a paper fold drawn on a flat canvas, Sonnet 5.5. https://t…
A pop-up city where each building is a paper fold drawn on a flat canvas, Sonnet 5.5. https://t.co/04cupOr4Tg
Why AI Agents Cheat | Eric Ho (Goodfire)
Speaker 1 | 00:00 - 00:20 A lot of people who aren't in the AI field are very surprised to realize that even the, you know, smartest researchers and scientists at the frontier don't understand their creations. It's like, oh, I thought we h…
I asked Secretary Duffy what Americans should be able to do, if we get this decade right, that…
I asked Secretary Duffy what Americans should be able to do, if we get this decade right, that they can't do today. Sean Duffy is the Secretary of Transportation, and he sat down with me at our Startup Industrial Base e…
I'm finding inter-agent communication in the chat stream increasingly irritating. Changed the v…
I'm finding inter-agent communication in the chat stream increasingly irritating. Changed the visibility in the OC harness to just a single line that can be expanded. Pretty sure others will follow. https://t.co/YUcQBzY…
[AINews] AMD buys World Labs for $8.2B, as Atlas solves sparse reconstruction problem for robotics, design and more
Congrats team!
Claude discovers a novel enzyme system \ Anthropic
In early results from our new life sciences research lab, Claude agents found an enzyme system whose function is still unknown.
Offloaded inference for real-world physical AI robotics
Robots are getting smarter, but how can their hardware match that growth? New Microsoft Research findings show that moving AI inference beyond the robot can improve task success, boost efficiency, and support more advanced physical AI work…
Xi Descends on DC
With Julian Gewirtz
At Box, we've been testing Sonnet 5.5 in early access on our complex work eval with the Box Age…
At Box, we've been testing Sonnet 5.5 in early access on our complex work eval with the Box Agent, and Sonnet 5.5 is another strong jump from Sonnet 5 on knowledge work with enterprise content. Across the board we saw a…
[AINews] The Future of Latent Space
A quiet day lets us discuss the work behind the scenes - now open for business!
Are you a Codex Original?
We’re collecting real stories of builders, tinkerers, researchers, and creators who are using Codex to do incredible things. If you want to be a part of the next chapter of the Codex Originals program, tell us more about your story and pro…
I wanted to know if I only hate perfume because I've never owned a nice one. So I researched a…
I wanted to know if I only hate perfume because I've never owned a nice one. So I researched a bunch of the fanciest and most well reviewed perfumes and bought a tiny sample bottle of each. I can now confidently confirm…
It continues to be notable how many AI researchers (who understand how AI actually works) belie…
It continues to be notable how many AI researchers (who understand how AI actually works) believe in neither AI doom nor massive uncontrollable acceleration, while non AI researchers (who have limited knowledge) have ve…
Foundries vs Navigators: Lowering the Cost of Science
Guest Post: In science, thinking has gotten cheap but doing has not. This asymmetry is reshaping how research companies operate, largely inconspicuously.
Introducing the Life Sciences Verification Program \ Anthropic
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Developing Enterprise Frontier Safeguards with our customers \ Anthropic
Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.
Improving synthesis prediction of small molecules at scale with RetroChimera
Custom-made molecules are advancing medicine, materials, and agriculture, but producing them is slow and expensive. A new Nature paper highlights RetroChimera, a predictive model that helps accelerate chemical synthesis, helping researcher…
Import AI 473: The US’s superintelligence strategy; human brain in a mouse skull; and machine hermeneutics
Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers. If you’d like to support this, please subscribe. Subscribe now RAND thinks the best AI strategy for the US is to keep al…
Re-Founding Incumbents for the AI Era with Sequence Holdings Co-Founder and CEO Michael Lee
Speaker 1 | 00:00 - 00:25 Every company on the planet has a celebrated persona. In a world where you believe that alpha comes from engineering and AI, you need to create a culture whereby the celebrated persona is the engineer, and that's…
I Went to Kyrgyzstan to Find China’s Silent Nomads
Representing Inner Mongolia in Kyrgyzstan
More details for the formal methods people -- what's happening is Claude is doing something lik…
More details for the formal methods people -- what's happening is Claude is doing something like: 1. Building a model of the program, targeting a tricky state machine or race-prone part of the code 2. Finding counter-ex…
New experts join Google’s AI & Economy team
We are expanding our AI & Economy team with world-class academic advisors, fellows, and core internal researchers.
How Chinese AI Radicalizes
Deepseek, Dario, and the boy who cried wolf
🔬 An Oscar, Two Asteroids, and the Algorithm in Your sklearn: John Platt on AI for Science
We talked to Google’s Oscar winning “Giganerd” about automating science, solving climate change, and how future generations can contribute to science in the age of superintelligent AI
I used Opus 5.5 to formally verify the Claude Agent SDK using Lean. A couple short prompts = 16…
I used Opus 5.5 to formally verify the Claude Agent SDK using Lean. A couple short prompts = 16 PRs fixing various bugs and race conditions. Video attached. TLA+ also works well. I sometimes combine Lean and TLA+ to loo…
Parallel cut research time and cost in half with GPT‑6 Astra
GPT‑6 Astra allowed Parallel’s agents to research and synthesize labor-market data in half the time and at half the cost vs. prior models.
Previewing the Model Hardware Standard \ Anthropic
Anthropic is opening a research preview of the Model Hardware Standard (MHS), a shared specification for AI agents to safely operate physical devices, to a first group of scientific research labs and advanced manufacturers.
Can Taiwan Build Drones? Part 2
This is Part 2 of our Taiwan drone series!