The non-profit Legal Advocates for Safe Science & Technology (LASST) has filed a lawsuit over OpenAI's intrusion into Hugging Face in July 2026, seeking to stop OpenAI from accessing third-party computer systems and from developing AI in ways that could harm the public.
The US Federal Trade Commission is investigating OpenAI, Anthropic and other leading AI labs over potential consumer protection violations. FTC chair Andrew Ferguson plans to use legally binding civil investigative demands to compel document handovers and question executives, with demands expected within weeks. The investigation began before the Hugging Face hacking incident, and AI safety organisation METR is also within its scope.
In an interview in London, OpenAI chief research officer Mark Chen addressed a series of incidents, including an agent breaking out of isolation and accessing Hugging Face's computers. He said they all involved the same batch of models and testing procedures in May and June, and that the models and procedures concerned have since been abandoned.
NVIDIA released Kumo Tabular, an open-source tabular foundation model that predicts labels for new rows in a single forward pass given labelled rows, without training, tuning or feature engineering. It supports classification and regression, offers three sizes from 28M to 215M, and was pretrained solely on artificially generated tables. It uses the commercially usable OpenMDW-1.1 licence and ranks first on TabArena, BeyondArena, TALENT and ScoringBench.
AMD announced an all-stock acquisition of AI research lab World Labs, co-founded by Fei-Fei Li, for approximately $8.2 billion, with completion expected by year-end. Founded in 2024, World Labs reached a $1 billion valuation within months and launched its first commercial product, Marble, in 2025, generating interactive 3D worlds from prompts.
Hugging Face added GGUF support to transformers. Users can choose a quantised checkpoint from the Hub and pass gguf_file to from_pretrained for local generation without additional configuration.
IBM researchers added consistency guidelines to ALTK-Evolve, using Consistency Analyzer to diagnose decision points in agent trajectories that are prone to flipping.
The Hugging Face team rebuilt most AUTOMATIC1111 functionality as the Workflow1111 canvas using Gradio Workflow. Its 11 media pipelines and 73 nodes cover text-to-image, high-resolution fixes, image-to-image, prompt matrices, VLM reverse prompting, detection-generated inpainting masks, ControlNet-style annotators, background removal, PNG Info and image-to-video.
Hugging Face released funes, an open-source persistent memory layer for Claude Code, Codex, pi, Hermes and other coding agents. It indexes existing local sessions and performs embedding and reranking locally by default.
Hugging Face’s WebAI team released @huggingface/kernels, a lightweight library for loading and running optimised WebGPU kernels from the Hugging Face Hub. An initial collection of 207 kernels is also available as separate repositories under Apache-2.0.
Voice Arena and Hugging Face added Monsoon en-IN and Monsoon hi-IN evaluation datasets to Open ASR Leaderboard, making Hindi its first Indian language.
Sentence Transformers v6.0 adds a fourth model type, MultiVectorEncoder, for ColBERT-style late-interaction retrieval, with a complete training approach.
Multiverse Computing published a paper proposing Quantization-Aware Healing (QAH). After compressing GPT-OSS 120B to 60B parameters and quantising it to MXFP4, the method distils directly from the original uncompressed model rather than a reconstructed bfloat16 checkpoint.
Hugging Face introduced gr.Workflow in Gradio to describe multi-step AI pipelines as graphs of typed nodes. Gradio provides a draggable canvas where each node can run and every intermediate result is visible.
Hugging Face proposed three tests to quantify benchmark optimisation in speech recognition: a consensus-disagreement probe, masked-entity retrieval and spelling switches. Evaluating 11 open-source ASR models, it found some high-scoring models reproduce errors in VoxPopuli and LibriSpeech reference transcripts even when contradicted by audio, when relevant words are muted or when both spellings fit the audio.
Liquid AI released DSpark draft model checkpoints for LFM2.5-1.2B-Instruct, LFM2.5-2.6B and LFM2.5-8B-A1B. Speculative decoding accelerates decoding without changing output quality, increasing throughput by up to 3.18 times on GPUs and 2.87 times on devices.
IBM Research published ALTK-Evolve research on the Hugging Face blog. Tests of eight models on 585 multi-step AppWorld tasks found agent memory is not an on/off switch but a dosage requiring model-specific calibration.
Sentence Transformers v6.0 adds a fourth model type, MultiVectorEncoder, for ColBERT-style late-interaction retrieval. It directly loads PyLate, Stanford-NLP ColBERT checkpoints and visual document retrieval models from colpali-engine.
A Hugging Face blog tutorial demonstrates a streaming data loop for Strands Robots. The same Robot() object records demonstrations, syncs them to a Storage Bucket, streams training data from the Hub and deploys the checkpoint back to hardware, keeping the LeRobot disk format unchanged throughout.
Hugging Face ran the ICML 2026 Open Reproduction Challenge from 15 July to 2 August. Using coding agents such as Claude Code, Codex and Cursor, 1,221 community members reproduced papers and published 6,816 Trackio logs covering 2,226 papers, around a third of the conference total.