A preprint is a signal, not a finished product or an independently confirmed result. We therefore track research papers separately and show the actual state of supporting evidence for each one.
268
published research events
Latest work
A significant claim from a single source is published only after further confirmation.
NADI 2026 is the seventh edition of the shared task focused on Arabic dialects and the second edition dedicated to speech processing. It includes five tasks: automatic speech recognition (ASR), dialect identification (SDID), text-to-speech synthesis (TTS), speech translation (SLT), and spoken language understanding (SLU). 21 teams from at least 13 countries participated…
Researchers from the MIT Senseable City Lab have published a book titled "How AI Sees the City: Urban Visual Intelligence" on the application of computer vision in urban planning. The book documents methods for detecting vehicles and their emissions, analyzing traffic and the movement of people in urban spaces, while also highlighting privacy risks and…
A research paper on arXiv presents the RefineICL method for tabular foundation models with in-situ representation refinement, where support labels drive updates to representations without changing parameters. The model achieves 0.93836 OVR-AUC on AMLB29 and 1644.8 Elo on TabArena, which is 31.4 Elo more than TabPFN-3.…
A research paper describes a semi-supervised federated ASR method combining online pseudo-labeling with server update stabilization. On nine of eleven tested pairs, it achieved an improvement of 20.8% (in-domain) and 10.0% (cross-domain) over previous methods.
The article argues against using AI-detection software in academic integrity. The NSW Education Standards Authority has banned schools from relying on these tools. An estimate shows that a 0.7% false positive rate would, in Australia, mean over 100 000 wrongly flagged tests per year. Turnitin…
A Gallup survey (37 countries) shows that 68% of Americans who use AI daily are worried about it, and only 36% believe it will be beneficial to the country. In other countries (Singapore 46% of daily users, China 35%), optimism is higher. The survey will be expanded to 140 countries.
404 Media (investigative tech journalism)
· and 1 more source
Original source ↗
Microsoft Research published a study on inference infrastructure for physical AI robots. They found that moving computation from onboard GPUs to the cloud/edge improves latency, accuracy, and autonomy, and reduces costs for mobile manipulation robots. Mapping/planning was up to 383% slower on small GPUs than on…
AI models achieve high accuracy in global weather forecasts, but fail at the regional scale, especially when predicting rapid hurricane intensification. The cause is a lack of detailed observational data outside coastal areas. Hurricane Polo intensified on 21 September 2026 from a tropical storm…
OpenAI announced MentalHealthBench, a benchmark designed to evaluate the helpfulness and safety of AI responses in realistic conversations focused on mental health. The methodology is based on data from experts.
The new SWE-Serve benchmark includes 53 production tasks from SGLang for evaluating AI agents in inference engineering. Among 11 models, the best configurations achieved 75 % pass@1. End-to-end tests rejected approximately a third of the patches that passed the other tests, exposing a gap between local development and…
The research team proposed a four-stage framework (SFT → PG-CoT → Dynamic → K-RL) to improve the generation of TCM formulations. It addresses three problems: a lack of auditable reasoning, missing patient monitoring, and violations of pharmacological rules. The Mistral-7B model outperformed GPT-5 in a zero-shot evaluation.
A scientific paper introduces the CMC framework to reduce memory overhead and speed up inference in large language models. The method compresses long contexts into context memory vectors aligned with the decoder. Experiments on four QA benchmarks: gains of up to 7.3 EM and 4.0 F1 points on SQuAD, 20%…
A study on arXiv tested whether synthetic personas in LLMs (GPT-4.1, Gemini in three versions) correctly predict which headlines an audience will click on. Using data from Upworthy Archive (399 statistically significant A/B tests), it found that querying the model directly without personas (Kendall τ = 0.361) outperforms persona-based…
A research paper proposes sheaf regularization to improve the stability of continuous clustering in the Decentralized SyncMap system. The method achieved the highest normalized mutual information (NMI) on 12 of 18 CGCP test graphs with two-state memory and 17 of 18 with dynamic memory. The algorithm adapts to…
A research team introduced Orthrus, an inference serving system that combines embedding and generative models in a single loop using heterogeneous batching. On four A100 GPUs, it achieved 1.28–4.52× higher throughput and up to 55.8% lower p99 latency compared with the baseline. The code has been released.
A research study evaluated five open language models on 3 600 numerical tasks with 8 600 transformed representations (fractions, percentages, scientific notation, unit conversions). Canonical accuracy reaches 0.969–0.996, but accuracy when the representation changes drops to 0.848–0.981. The model Mistral…
The study evaluates diachronic word embeddings for Sanskrit — an ancient language with limited data resources and high morphological complexity. It collected 2.7M tokens from four canonical periods and applied a neural sandhi splitter and a lemmatizer. Of the 21 semantic shifts tested, 19 moved in the correct direction…
An empirical study of 12 MoE models (9 architectures) demonstrated that uniform pruning—retaining only 2/3 of the selected experts—achieves 98.8 % of the original performance with a 1.2–1.7x inference speedup. The simple strategy is as effective as more complex dynamic methods under conservative budgets.
AI Radar monitors Czech and international sources every day, looking for changes that truly deserve attention.
MonitorsOfficial AI company blogs, specialist media, and research sources.
Selects and combinesFilters out information noise and combines articles about the same change into a single event.
Summarizes and explainsExplains significant events in English: what happened, why it matters and where the information comes from.
The result is a quick overview of what has actually changed in the AI world, rather than another stream of articles.
Use the CS/EN switch to read the same Radar in Czech or English. English content is published after its translation has been checked, so new and older items may appear later.
Everything you need to navigate the AI world
Today’s briefingThe “What is worth attention” selection sits beside Live · AI Flash, followed by research and links to other Radar sections. On mobile, these blocks appear one below another.
AI FlashAn ongoing feed of brief updates with an evidence status. Links lead to a Radar detail page when one is ready, otherwise to the original source. You can also find reset and outage histories here.
Practical applicationsWhat new tools and features can do, what you can try and what their actual impact could be.
Model selectionModel comparison by type of work, capabilities, price and speed.
Research and archiveA separate research overview, topic search and older events by date.
One event, everything that matters
Each row represents one event — not one article. At a glance, you can see its significance, credibility and main point.
Illustrative example, not a current news item.
Importance: ▮▮▮ majorOpenAIModels✓ 6
Agent mode is available to all paying users
Until now, the mode was available only on the highest plan; it is now available on all paid tiers without a waitlist.
▮▮▮ major · ▮▮ important · ▮ we're tracking = how significant the change is✓ 6 = six independent publishers, not the number of articles or feeds✓ official = a clear release, law or incident is substantiated by the relevant authority1 source = no independent confirmation yetbold = who is behind the changegray text = a brief summary of what happened
The detail page contains a fuller summary, its significance and original sources. Practical impact appears in the detail and the For individuals and For businesses views. An AI Flash item reaches the main selection only after it has been expanded and meets the publication rules.
The same news, two practical uses
We first summarize each event in the same way for everyone. Based on those same facts, we then explain what the change means for your own use and what it could mean for how a company operates.
For individualsWhat you can use or try, how the change can help you at work and what to watch out for.
For businessesWhat impact the change could have on processes, costs, risks and other business decisions.
Today’s briefing is the same for everyone. Pages
For individuals and For businesses
can be found in the main navigation — they select only events relevant to the given use case.