AINewsnow

How to automatically find the batch size when using Accelerate with FSDP2? [D]

This story is from 2026-09-14. It is preserved in the archive; the latest stories are on the live feed.

Hi, For single-GPU training, I’m using Hugging Face SFTTrainer with auto_find_batch_size=True, which automatically reduces the batch size after a CUDA OOM until it finds a batch size that works. I would like to have similar behavior when training on multiple GPUs on a single node using accelerate l…

Read the full story at r/MachineLearning ↗

Timeline · 1 report

  1. 2026-09-14 17:24 · r/MachineLearning
    How to automatically find the batch size when using Accelerate with FSDP2? [D]

More stories

  1. We’re Not Losing Control of A.I. We’re Giving It Away. — New York Times AI
  2. Deploy Hugging Face models on Amazon SageMaker AI with coding agents — AWS Machine Learning Blog
  3. A quick Minimax H3 news round-up - 17th September 2026 — r/comfyui
  4. Hugging Face Hack Shows Humans Can Keep AI In Check — AI Now Institute
  5. ‘Godfather of AI’ Geoffrey Hinton warns humans running out of time to control Artificial Intelligence – ‘maybe a year’ — Mint AI
  6. Your AI agents are isolated. Your infrastructure isn’t — InfoWorld AI
  7. ‘Jailbreak-like...’: AI's ‘unexpected’ behaviour mounts concerns, OpenAI's 'rogue agents probed' Hugging Face — Mint AI
  8. What is actually going on with all the recent AI safety / “rogue agent” stories? — r/ArtificialInteligence

Get the daily brief of stories like this at 6:30 every morning →