Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:OpenRLHF OpenRLHF SFT Training
- Workflow:Mlc ai Mlc llm REST API Serving
- Workflow:Scikit learn contrib Imbalanced learn Imbalanced Model Evaluation
- Workflow:Open compass VLMEvalKit Video Benchmark Evaluation
- Workflow:Explodinggradients Ragas RAG Evaluation
- Workflow:Elevenlabs Elevenlabs python Realtime TTS Streaming
- Workflow:Farama Foundation Gymnasium Custom Environment Creation
- Workflow:Zai org CogVideo SAT Finetuning
- Workflow:Deepspeedai DeepSpeed AutoTP Training
- Workflow:Mistralai Client python Text Embeddings
Principles
- Principle:Volcengine Verl Multi Turn Rollout
- Principle:ArroyoSystems Arroyo Pipeline Submission
- Principle:DistrictDataLabs Yellowbrick Visual Pipeline Integration
- Principle:Puppeteer Puppeteer Browser Version Verification
- Principle:Google deepmind Mujoco Offscreen Rendering
- Principle:Tensorflow Tfjs Weight Regularization
- Principle:Bentoml BentoML Distributed Deployment Configuration
- Principle:Ggml org Llama cpp StreamingJSONParsing
- Principle:Turboderp org Exllamav2 Bit Allocation Optimization
- Principle:Datahub project Datahub Client Authentication
Implementations
- Implementation:Explodinggradients Ragas FactualCorrectness Metric
- Implementation:CrewAIInc CrewAI RAG PostgreSQL Loader
- Implementation:Webdriverio Webdriverio Bidi LocalTypes
- Implementation:Turboderp org Exllamav2 Ext Cache
- Implementation:Evidentlyai Evidently SDK Panels
- Implementation:Cleanlab Cleanlab Rank Classes By Label Quality
- Implementation:EvolvingLMMs Lab Lmms eval VDC Video Captioning Utils
- Implementation:Deepset ai Haystack PyPDFToDocument
- Implementation:Datajuicer Data juicer ImageNSFWFilter
- Implementation:Arize ai Phoenix LangChain Adapter
Heuristics
- Heuristic:Deepset ai Haystack Pipeline Max Runs Safety Limit
- Heuristic:Duckdb Duckdb PR Submission Strategy
- Heuristic:Lm sys FastChat Vicuna SFT Training Hyperparameters
- Heuristic:Facebookresearch Habitat lab Resume State Config Override
- Heuristic:AUTOMATIC1111 Stable diffusion webui VRAM Management Strategies
- Heuristic:Cleanlab Cleanlab Label Quality Scoring Method Selection
- Heuristic:Ray project Ray Serve Concurrency And Backpressure
- Heuristic:Allenai Open instruct Logprob Clamping
- Heuristic:Axolotl ai cloud Axolotl Memory Optimization Tips
- Heuristic:Shiyu coder Kronos Gradient Clipping Strategy
Environments
- Environment:Google research Deduplicate text datasets Python HuggingFace Environment
- Environment:SeldonIO Seldon core Kubernetes Cluster Environment
- Environment:Pyro ppl Pyro CUDA GPU Acceleration
- Environment:SeleniumHQ Selenium Contributor Development Environment
- Environment:Pyro ppl Pyro Python PyTorch Core
- Environment:Volcengine Verl CUDA GPU Environment
- Environment:Mlc ai Mlc llm Metal macOS iOS Environment
- Environment:OpenBMB UltraFeedback OpenAI API Environment
- Environment:Huggingface Open r1 CUDA Environment
- Environment:OWASP Www project top 10 for large language model applications Pre Commit Hooks Environment