Skip to content
#

gradient-checkpointing

Here are 9 public repositories matching this topic...

End-to-end fine-tuning of Hugging Face models using LoRA, QLoRA, quantization, and PEFT techniques. Optimized for low-memory with efficient model deployment

  • Updated Dec 27, 2025
  • Jupyter Notebook

Fine-tune NVIDIA Isaac GR00T N1.7 (3B VLA) on a single 24 GB RTX 4090 without root: patches, install script, MuJoCo evaluation, open datasets, 1800 evaluated attempts as JSON

  • Updated Sep 8, 2026
  • Python

Classify documents longer than your transformer's context window: overlapping windows, cross-window aggregation (max / mean / top-k / log-mean-exp), and window-level gradient checkpointing that cuts training memory. Encoder-agnostic and task-agnostic, with measured speed and accuracy trade-offs.

  • Updated Sep 23, 2026
  • Python

Add this topic to your repo

To associate your repository with the gradient-checkpointing topic, visit your repo's landing page and select "manage topics."

Learn more