A fast, efficient universal vector embedding utility package.
-
Updated
Aug 3, 2023 - Python
A fast, efficient universal vector embedding utility package.
Large Context Attention
Implementation of a memory efficient multi-head attention as proposed in the paper, "Self-attention Does Not Need O(n²) Memory"
APOLLO: SGD-like Memory, AdamW-level Performance; MLSys'25 Oustanding Paper Honorable Mention
VGGT-X: When VGGT Meets Dense Novel View Synthesis
[CVPR 2025 Highlight] The official CLIP training codebase of Inf-CL: "Breaking the Memory Barrier: Near Infinite Batch Size Scaling for Contrastive Loss". A super memory-efficiency CLIP training scheme.
Easy Parallel Library (EPL) is a general and efficient deep learning framework for distributed model training.
Train Dense Passage Retriever (DPR) with a single GPU
Code for: "Social-Implicit: Rethinking Trajectory Prediction Evaluation and The Effectiveness of Implicit Maximum Likelihood Estimation" Accepted @ ECCV2022
[ACL 2023] The official implementation of "CAME: Confidence-guided Adaptive Memory Optimization"
Efficient Multimodal Foundation Model Adaptation for Recommendation
A family of highly efficient, lightweight yet powerful optimizers.
Face search engine: 100% recall at 96x less memory than HNSW. Register photos, identify people from camera/video/images. AVX-512 + Rust. No GPU needed.
Tiered optimizer state allocation for memory efficient MoE training. Cuts optimizer memory by 97.4%, outperforming AdamW/Muon/Lion while fitting a 6.78B MoE on a single 40GB GPU.
Lowering PyTorch's Memory Consumption for Selective Differentiation
Scalable Parameter and Memory Efficient Pretraining for LLM: Recent Algorithmic Advances and Benchmarking
The official implementation of Memory-efficient DQN algorithm.
A fast and memory efficient way to load large CSV files (Timeseries data) in Pandas.
Lion (EvoLved Sign Momentum) optimizer utilizing sign operations for memory-efficient uniform gradient updates.
Lion (EvoLved Sign Momentum) optimizer utilizing sign operations for memory-efficient uniform gradient updates.
To associate your repository with the memory-efficient topic, visit your repo's landing page and select "manage topics."