Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Ggml org Ggml Backend Accelerated Computation
- Workflow:Togethercomputer Together python Chat Completion
- Workflow:Haifengl Smile Nearest Neighbor Search
- Workflow:Norrrrrrr lyn WAInjectBench LLaVA Finetuning
- Workflow:Mlc ai Mlc llm Model Compilation
- Workflow:Scikit learn contrib Imbalanced learn SMOTE Resampling Pipeline
- Workflow:Elevenlabs Elevenlabs python Speech to Text Transcription
- Workflow:Rapidsai Cuml Multi GPU Distributed ML
- Workflow:Zai org CogVideo Diffusers LoRA Finetuning
- Workflow:Unstructured IO Unstructured Chunking And Embedding
Principles
- Principle:Huggingface Datatrove Document Tokenization
- Principle:Lucidrains X transformers Hybrid Discrete Continuous Tokens
- Principle:FlagOpen FlagEmbedding Finetuned Reranker Validation
- Principle:Mlfoundations Open flamingo WebDataset Data Pipeline
- Principle:CARLA simulator Carla Pedestrian Navigation Mesh
- Principle:DistrictDataLabs Yellowbrick Classification Report Visualization
- Principle:Scikit learn Scikit learn Bayesian Regression
- Principle:Scikit learn Scikit learn Learning Curve Analysis
- Principle:Huggingface Diffusers Single File Loading
- Principle:Princeton nlp Tree of thought llm Usage Tracking
Implementations
- Implementation:ARISE Initiative Robosuite Demo Sensor Corruption
- Implementation:Confident ai Deepeval KeyFileHandler
- Implementation:Tensorflow Tfjs Merge Test
- Implementation:Vespa engine Vespa Vespajlib ABI Spec
- Implementation:Unslothai Unsloth Evaluate OCR Model
- Implementation:ArroyoSystems Arroyo Create Connection Table
- Implementation:Pola rs Polars Buffer
- Implementation:Mit han lab Llm awq Pseudo quantize model weight
- Implementation:Huggingface Diffusers ControlNetModel Forward
- Implementation:OpenHands OpenHands SaasNestedConversationManager Maybe Start Agent Loop
Heuristics
- Heuristic:Facebookresearch Habitat lab Warning Deprecated Legacy UI System
- Heuristic:Snorkel team Snorkel Precision Init Prior
- Heuristic:Datahub project Datahub Docker Memory Preflight
- Heuristic:MarketSquare Robotframework browser Shared Node Process For Parallel
- Heuristic:Haosulab ManiSkill GPU Memory Buffer Tuning
- Heuristic:ThreeSR Awesome Inference Time Scaling Empty Venue Default Tip
- Heuristic:DataExpert io Data engineer handbook Docker Volume Persistence Management
- Heuristic:Haosulab ManiSkill Physics Solver Tuning
- Heuristic:BerriAI Litellm Streaming Loop Detection
- Heuristic:Unslothai Unsloth Gradient Accumulation Accuracy
Environments
- Environment:OWASP Www project top 10 for large language model applications Pydantic Invoice Agent Runtime
- Environment:AUTOMATIC1111 Stable diffusion webui Python And PyTorch Runtime
- Environment:Unstructured IO Unstructured GitHub Actions
- Environment:Allenai Open instruct CUDA GPU Training
- Environment:Datajuicer Data juicer Ray Cluster Environment
- Environment:Huggingface Alignment handbook Python TRL
- Environment:Ollama Ollama CGo Runtime
- Environment:Ray project Ray CI Build Matrix Environment
- Environment:Openai CLIP Python Dependencies
- Environment:Facebookresearch Audiocraft Python PyTorch CUDA Environment