Skip to content

News: selected

Automating coherent long-form video generation

Google Research published research on long-form video generation, proposing an AI video co-director multi-agent orchestration framework built on Gemini and Veo to plan visual continuity across multi-shot narratives.

Google Research

Offloaded inference for real-world physical AI robotics

Microsoft Research systematically measured mobile manipulation robot workloads and found that offloading physical AI inference from onboard GPUs to edge or cloud GPUs can improve task success, support larger models and extend battery life.

Microsoft Research

Your Agent Aced the Task. Will It Do It Again?

IBM researchers added consistency guidelines to ALTK-Evolve, using Consistency Analyzer to diagnose decision points in agent trajectories that are prone to flipping.

Hugging Face Blog

A connectomics milestone: Mapping the complete male fruit fly brain

Google Research partnered with HHMI Janelia, the University of Cambridge and others to publish in Cell a complete connectome of a male fruit fly’s brain and central nervous system. It contains over 166,000 neurons and 125 million synaptic connections, making it the largest brain map to date by neuron count.

Google Research

  1. Planetary prediction engine: Automating global models via Earth AI

    Google Research introduced the experimental Planetary Prediction Engine (PPE) under Google Earth AI. From a natural language query, it autonomously discovers geospatial data, engineers features, trains and evaluates models and produces reports, compressing weeks of manual data engineering into minutes.

    Google ResearchPapers

Measuring benchmark optimization in speech recognition

Hugging Face proposed three tests to quantify benchmark optimisation in speech recognition: a consensus-disagreement probe, masked-entity retrieval and spelling switches. Evaluating 11 open-source ASR models, it found some high-scoring models reproduce errors in VoxPopuli and LibriSpeech reference transcripts even when contradicted by audio, when relevant words are muted or when both spellings fit the audio.

Hugging Face Blog

What We Learned by Reproducing 2,200 papers from ICML

Hugging Face ran the ICML 2026 Open Reproduction Challenge from 15 July to 2 August. Using coding agents such as Claude Code, Codex and Cursor, 1,221 community members reproduced papers and published 6,816 Trackio logs covering 2,226 papers, around a third of the conference total.

Hugging Face Blog

  1. Science One Framework: A verifiable autonomous research framework via Chain-of-Evidence

    Google Research released the Chain-of-Evidence (CoE) verifiability framework, implemented in a Science One Framework prototype, with automated CoE Audit metrics.

    Google ResearchPapers
  1. Towards a quantum computer that learns from its errors

    Google Research published a study in Nature proposing a reinforcement learning framework in which agents continuously learn from quantum error-correction detection events, dynamically adjusting thousands of control parameters during computation to counter drift.

    Google ResearchPapers

SensorFM: Towards a general intelligence and interface for wearable health data

Google Research and Google DeepMind proposed SensorFM, a large sensor foundation model learning directly from unlabelled wearable data. Pretraining uses over 1 trillion minutes of multimodal sensor signals from 5 million consenting participants across more than 100 countries and over 20 Fitbit and Pixel Watch devices.

Google Research

Measuring the impact of learning with AI in Sierra Leone and beyond

Google DeepMind published results from a preregistered randomised controlled trial in Sierra Leone. Over eight weeks, students using Guided Learning improved maths scores by 0.258 standard deviations over controls, equivalent to around 1.2–1.7 years of conventional progress.

Google DeepMind