Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Rapidsai Cuml GPU Clustering
- Workflow:OpenHands OpenHands SaaS Server Assembly
- Workflow:Datajuicer Data juicer Dataset Quality Analysis
- Workflow:NVIDIA TransformerEngine Comm GEMM Overlap Training
- Workflow:Puppeteer Puppeteer Page Screenshot Capture
- Workflow:Zai org CogVideo Diffusers Image to Video Inference
- Workflow:Apache Hudi Docker Demo Setup
- Workflow:Scikit learn Scikit learn Supervised Classification
- Workflow:Deepseek ai Janus Autoregressive Image Generation
- Workflow:Danijar Dreamerv3 Evaluation Only
Principles
- Principle:Pyro ppl Pyro Causal Effect Estimation
- Principle:Huggingface Datatrove Core Data Structures
- Principle:Sail sg LongSpec Benchmark Data Preparation
- Principle:Microsoft Agent framework Framework Installation
- Principle:Langgenius Dify APIContracts
- Principle:Tensorflow Serving HTTP Server Interface
- Principle:Run llama Llama index QA Pair Generation
- Principle:Confident ai Deepeval Synthetic Data Synthesis
- Principle:Haosulab ManiSkill Demonstration Data Acquisition
- Principle:Ggml org Llama cpp Legacy Model Conversion
Implementations
- Implementation:Hiyouga LLaMA Factory VLLM Engine
- Implementation:Astronomer Astronomer cosmos Cluster Policy
- Implementation:Speechbrain Speechbrain Train CommonVoice Seq2Seq
- Implementation:Kubeflow Kubeflow Manifests E2E Testing
- Implementation:Lance format Lance LegacyDecoder
- Implementation:FlagOpen FlagEmbedding LLM Embedder Retrieval Args
- Implementation:Microsoft Onnxruntime CPU TrainingSplit
- Implementation:TobikoData Sqlmesh SelectEnvironment
- Implementation:Microsoft Semantic kernel QualityCheck NLP Server
- Implementation:Langchain ai Langchain BaseChatModel Stream
Heuristics
- Heuristic:Junyanz Pytorch CycleGAN and pix2pix Identity Loss Color Preservation
- Heuristic:Hiyouga LLaMA Factory Mixed Precision Training Tips
- Heuristic:Apache Druid Sampler Limitations And Workarounds
- Heuristic:Datahub project Datahub Gradle Formatting Over Direct Tools
- Heuristic:OpenGVLab InternVL Packed Training Buffer Management
- Heuristic:DistrictDataLabs Yellowbrick Scikit Learn API Compatibility
- Heuristic:Huggingface Diffusers VAE Scaling Factors
- Heuristic:Intel Ipex llm LoRA Target All Linear Layers
- Heuristic:Huggingface Open r1 vLLM GPU Allocation
- Heuristic:Huggingface Datatrove FineWeb Filter Pipeline Order
Environments
- Environment:Testtimescaling Testtimescaling github io GitHub Actions Runner
- Environment:Tencent Ncnn Vulkan Environment
- Environment:Huggingface Diffusers Quantization Environment
- Environment:ARISE Initiative Robomimic HuggingFace Hub Dependencies
- Environment:Anthropics Anthropic sdk python AWS Bedrock Environment
- Environment:Bitsandbytes foundation Bitsandbytes Build From Source Environment
- Environment:DataTalksClub Data engineering zoomcamp Dbt DuckDB Environment
- Environment:Pytorch Serve vLLM Engine Environment
- Environment:VainF Torch Pruning PyTorch Python Core
- Environment:MarketSquare Robotframework browser Docker Container