Skip to content

#Hugging Face

0 today

Jul 21

Tuesday
  1. Grabette: an open system to record robot-manipulation data

    Hugging Face released Grabette, an open-source system recording manipulation demonstrations with a handheld gripper and two cameras, producing robot-ready datasets without robots or teleoperation equipment. The handheld hardware costs about €490 in materials, with the accompanying motorised Gripette gripper around €120. Hardware CAD, Raspberry Pi collection software and browser-based processing are all open-source.

Jul 16

Thursday
  1. Newer Models, Same Advantage

    DharmaOCR scored 0.925 on a Portuguese OCR benchmark, ahead of Mistral OCR4's 0.798 and Unlimited-OCR's 0.7587. It specialised through two-stage training: supervised fine-tuning on Portuguese corpora, then DPO to stabilise inference. The author argues that concentrating parameters on one language remains a structural advantage despite emerging architectures.

  2. Security incident disclosure — July 2026

    Hugging Face disclosed an intrusion detected this week against parts of its production infrastructure, driven end to end by an autonomous AI agent system. Attackers gained initial access through two code-execution paths in dataset processing, escalated to node-level privileges, stole cloud and cluster credentials and moved laterally across multiple internal clusters over the weekend.

Jul 15

Wednesday

Jul 10

Friday

Jul 8

Wednesday

Jul 7

Tuesday

Jul 6

Monday
  1. PRX Part 4: Our Data Strategy

    Hugging Face's fourth PRX instalment details its data strategy: mixing public and internal pretraining datasets, regenerating long image captions with a VLM and converting them for streaming training. It uses Lance for construction and filtering and MDS for streaming. After switching to Qwen3-VL, text latents are computed during training, with measured throughput loss of around 3–4%, or roughly one extra day for 30 days of training.

Jul 2

Thursday

Jul 1

Wednesday

Jun 30

Tuesday

Jun 11

Thursday

Mar 24

Tuesday
  1. Speaking of Voxtral

    Mistral AI released its first text-to-speech model, Voxtral TTS, with 4B parameters and support for nine languages: English, French, German, Spanish, Dutch, Portuguese, Italian, Hindi and Arabic. It is available through the API and Mistral Studio at $0.016 per 1k characters.

Feb 5

Thursday

Jul 15

Tuesday
  1. Voxtral

    Mistral AI released Voxtral speech-understanding models in 24B and 3B versions, both under Apache 2.0 and available via its API. With a 32k-token context, they handle up to 30 minutes of transcription or 40 minutes of audio understanding, including built-in question answering and summaries, automatic multilingual detection and voice-triggered function calls, inheriting Mistral Small 3.1's text capabilities.

Jun 10

Tuesday

May 21

Wednesday

Mar 17

Monday