Hugging Face and SkyPilot released an integration allowing a Hugging Face Bucket or any model, dataset or Space repository to be mounted into SkyPilot tasks with an hf:// URL and an existing HF_TOKEN, running across more than 20 clouds, Kubernetes, Slurm and local environments.
Hugging Face and Cerebras jointly demonstrated a real-time speech-to-speech pipeline, accelerating Gemma 4 31B inference with Cerebras and combining Nvidia Parakeet speech recognition with Alibaba Qwen3TTS synthesis.
Google DeepMind integrated computer use as a built-in tool in Gemini 3.5 Flash. Previously, it was available only as the standalone Gemini 2.5 computer use model.
Mistral AI added new Connectors capabilities. Admin controls for workspace- or organisation-level access and individual tool toggles are generally available, as are API keys scoped to connectors.
Google launched a public preview of Agentic RAG-powered cross-corpus retrieval on Gemini Enterprise Agent Platform. Roles including Orchestrator, Planner, Query Rewriter, Search Fanout and Sufficient Context Agent collaborate on multi-source, multi-hop queries.
Google Research open-sourced its hydrological modelling framework on GitHub under Apache 2.0, enabling national weather and hydrology agencies to integrate AI flood forecasting. The Python package uses PyTorch and an LSTM architecture, can train or fine-tune on Caravan data, and includes interactive tutorial notebooks and videos.
Mistral released an AI stack for industrial engineering at AI Now Summit 2026, partnering with Airbus, BMW and ASML to optimise design, simulation and production while retaining control over proprietary data and IP.
Mistral AI released the public preview of Search Toolkit, a composable framework for production search pipelines in AI applications. It unifies ingestion, retrieval and evaluation through shared interfaces, is open-source and deploys in cloud, local or edge environments.
Mistral released Connectors in Studio. Built-in connectors and custom MCP are available through API/SDK for all model and agent calls, currently in public preview.
Google released Empirical Research Assistance (ERA), a research tool using Gemini to write and optimise scientific code. Its paper was published in Nature today, and the tool is available to scientists worldwide as part of Gemini for Science.
Google DeepMind added Street View grounding to experimental prototype Project Genie. Users select a real US location with a Maps pin, choose a style and describe a character, then Genie creates an interactive world whose starting location is grounded in real imagery.
Google announced the expansion of content transparency and verification tools to Search, Gemini, Chrome, Pixel and Cloud. SynthID has watermarked over 100 billion images and videos and 60,000 years of audio. SynthID verification in the Gemini app has been used 50 million times and will reach Search and Chrome in the coming weeks.
Mistral AI released Workflows in public preview, positioning it as an enterprise AI orchestration layer with durable execution, observability and fault tolerance to take AI workflows from proof of concept to production.
Google Research introduced a new method in Google Photos’ Auto frame to change camera viewpoints after a photo is taken. An internal 3D point-map estimation model reconstructs the scene and focal length, then a generative latent diffusion model fills gaps exposed by the new view. It automatically detects faces and wide-angle distortion to suggest ideal framing. The feature already applies automatically to photos containing people, with reframed versions available among Auto frame candidates.
Google DeepMind unveiled an experimental Gemini-powered AI pointer that understands not only what it points at but what it means to the user. The team proposed four interaction principles: avoiding workflow interruption, capturing nearby visual and semantic context, supporting natural shorthand such as “this” and “that”, and turning pixels into actionable entities such as places, dates and objects.
Google released Vibe Coding XR, combining Gemini with the XR Blocks framework based on WebXR, three.js and LiteRT.js to turn natural language prompts directly into physically aware Android XR apps, reportedly in under 60 seconds.
Mistral AI launched Forge, a system for enterprises to build frontier-class AI models on proprietary knowledge, supporting pre-training, post-training and reinforcement learning. It handles dense and MoE architectures and, when needed, multimodal inputs, with training and governance on companies' own infrastructure.