Skip to content

AI Flash · live feed

What is happening in AI right now

A timeline of everything that has passed through the radar — newest first, with an evidence status for each item.

9 new news items today 1141 across the entire feed ↗ details, original sources and our own analyses

Continuously updated

Latest items

1141 items in the selected feed

How to read evidence statuses ↓

Friday, July 24

13 items
Tools and apps only one source so far

Sakana claims that Fugu Ultra v1.1 outperforms Fable 5 without including it in the model pool

Sakana released Fugu Ultra v1.1, an update to the model router. It claims improvements of up to 7.9 points and that the system outperforms Fable 5, even though Fable 5 is not included in the router. Price: $5/$30 for millions of tokens. All figures come from Sakana; independent verification is lacking. An endpoint compatible with Claude Code has been added.

The Decoder (daily AI news) Original source ↗
Robotics only one source so far

AI-powered robotics in Europe: demonstrations and a strategic initiative

An upcoming meeting in Europe brings together policymakers, industry and academia to discuss European leadership in AI-powered robotics. euROBIN (a Horizon Europe network) is preparing an event with demonstrations of 20–30 advanced robots and a discussion of a joint European initiative for physical AI (robots that perceive, reason and act).…

European Commission — Shaping Europe's digital future (AI Act, EU regulation) Original source ↗
Marketing only one source so far

Young people are skeptical of AI marketing

Although young people use chatbots, 37 % of teenagers mock AI content and more than half fear disinformation and deepfakes. More than half of the generation is concerned about the negative impacts of AI on society, including the loss of jobs and critical thinking. Young people perceive AI marketing by adults as pressure.

Wired — AI section Original source ↗
Business and investment only one source so far

NVIDIA and KAIST will establish a joint AI lab focused on agentic AI

NVIDIA and South Korean university KAIST have established a joint AI research lab in Seoul focused on agentic AI. It is the first partnership of its kind between a Korean university and a global technology company. At the same time, collaboration between SK and NVIDIA on developing memory for AI platforms is expanding.

NVIDIA Newsroom (press releases) Original source ↗
New models only one source so far

MKB: A unified scientific multimodal foundation model

MKB (Monkey King Bang), a new model from the Shanghai Academy of AI for Science, unifies scientific research across six fields: DNA, RNA, proteins, small molecules, geophysics and medical images. It supports native outputs for biological sequences, molecular strings and meteorological fields. The model is built on…

arXiv cs.LG (Machine Learning) Original source ↗
AI agents only one source so far

GPT-5.5 Pro autonomously generates proofs of the sum-product conjecture

An arXiv preprint describes an autonomous agent based on GPT-5.5 Pro that generated mathematical proofs disproving the sum-product conjecture over the real numbers in 7 of 8 attempts, each using an average of 132.4 thousand reasoning tokens. The authors released the code and proofs for reproduction.

arXiv cs.AI (Artificial Intelligence) Original source ↗
AI agents only one source so far

AINTMA: Autonomous Test Management Architecture with Multi-Agent AI

A research paper presents AINTMA, a system of six AI agents for automated testing. It integrates LLMs for quality reports, reinforcement learning for test prioritization, and zero-trust security. It claims 88.4% accuracy (vs. 51.2% chance), a 43% reduction in testing time, and a reduction in defects from 8.3% to 2.1%.

arXiv cs.AI (Artificial Intelligence) Original source ↗
Security only one source so far

Safety safeguards (guardrails) in AI models complicate work for offensive cybersecurity researchers

According to TechCrunch, restrictions (guardrails) on models from Anthropic and OpenAI intended to prevent misuse for cyberattacks also hold back legitimate security researchers. The companies offer vetting programs with less restrictive limits, but researchers complain about inconsistent model behavior and bypass them with open-source…

Tools and apps only one source so far

Configuring security for coding assistants in Amazon Bedrock

A guide to configuring guardrails (safety measures) in Amazon Bedrock for AI assistants that generate code. It addresses performance and cost issues when deploying to multiple developers, with a case study involving 15 developers and throttling errors.

AWS Machine Learning Blog Original source ↗

Thursday, July 23

11 items
Security only one source so far

Judge identifies AI-generated errors in a court transcript

In a decision dated 23 July, Judge Paul Felix noted that a court transcript contained errors resembling generative AI output. The warning addresses court reporters about the risks of using AI tools in legal documents without verification.

404 Media (investigative tech journalism) Original source ↗
Tools and apps clearly official source

GitHub Mobile: Fixing failing tests through the Copilot agent

GitHub is adding the ability to ask the Copilot AI agent in its mobile app to investigate and fix failing GitHub Actions checks. The agent analyzes the error, creates a pull request with a proposed fix and assigns a developer to review it. The feature is available on iOS and Android.

GitHub Changelog (Copilot and AI features) Original source ↗
Tools and apps confirmed by 3 independent sources

Claude gets voice mode with Opus and Sonnet models and connections to Gmail, Calendar and Slack

Anthropic has expanded voice mode in Claude, previously based only on the Haiku model, to include the more powerful Opus and Sonnet. It now connects to Gmail, Google Calendar, Slack, Canva and Notion and can send an email or modify a meeting on behalf of the user through voice commands.

The Verge AI · and 2 more sources Details and sources →
AI agents only one source so far

An AI trading assistant: Jefferies automates trading desk analysis

Jefferies built an AI agent-based trading assistant on AWS that provides traders with real-time analysis of data from multiple sources using natural language, without coding. The solution uses Anthropic's Claude, Amazon Bedrock, Knowledge Bases and MCP to connect to market data. The assistant accelerated data exploration from days…

AWS Machine Learning Blog Original source ↗
AI agents only one source so far

Detecting silent AI agent failures in Amazon Bedrock

Amazon Bedrock AgentCore is an optimization feature that detects silent AI agent failures — errors without error signals, such as incorrectly performed actions or omitted checks. The tool analyzes traces using a taxonomy of 11 failure categories and groups them into patterns. According to Amazon, it helps distinguish…

AWS Machine Learning Blog Original source ↗