Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:ArroyoSystems Arroyo Local Pipeline Execution
- Workflow:Apache Kafka Docker Image Release
- Workflow:Google research Deduplicate text datasets Suffix array querying
- Workflow:Dotnet Machinelearning ONNX Model Scoring
- Workflow:Scikit learn contrib Imbalanced learn SMOTE Resampling Pipeline
- Workflow:Ollama Ollama Model Registry Operations
- Workflow:Teamcapybara Capybara Element Finding And Interaction
- Workflow:Huggingface Diffusers LoRA Finetuning
- Workflow:Kornia Kornia ONNX Model Pipeline
- Workflow:Huggingface Transformers PEFT Adapter Integration
Principles
- Principle:Elevenlabs Elevenlabs python Client Initialization
- Principle:Confident ai Deepeval Offline Trace Evaluation
- Principle:Huggingface Optimum Task and Model Resolution
- Principle:Cypress io Cypress Project Configuration
- Principle:Snorkel team Snorkel Label Quality Evaluation
- Principle:Apache Paimon Atomic Commit
- Principle:Microsoft Semantic kernel Vector Store Data Model
- Principle:Ollama Ollama Sampling Strategy
- Principle:Trailofbits Fickling ML Allowlist Unpickling
- Principle:Langchain ai Langgraph Graph Configuration Management
Implementations
- Implementation:Google deepmind Mujoco Engine Island
- Implementation:Trailofbits Fickling Get Stats
- Implementation:Open compass VLMEvalKit MLVU
- Implementation:FlagOpen FlagEmbedding Matryoshka Compensation Data
- Implementation:ArroyoSystems Arroyo Redis Connector
- Implementation:Microsoft Playwright Firefox Protocol Types
- Implementation:Protectai Llm guard TokenLimit
- Implementation:Datahub project Datahub ProtobufExtensionFieldVisitor
- Implementation:CrewAIInc CrewAI RAG Docs Site Loader
- Implementation:InternLM Lmdeploy Gemm SmemCopy
Heuristics
- Heuristic:Langgenius Dify Celery Queue Separation
- Heuristic:Microsoft Playwright Timeout Configuration Tips
- Heuristic:Onnx Onnx Opset Version Selection
- Heuristic:Bitsandbytes foundation Bitsandbytes Compute Dtype Mismatch Warning
- Heuristic:Risingwavelabs Risingwave Stream Chunk Sizing
- Heuristic:Duckdb Duckdb Test Development Guidelines
- Heuristic:Huggingface Transformers Gradient Checkpointing Memory Tradeoff
- Heuristic:Junyanz Pytorch CycleGAN and pix2pix Identity Loss Color Preservation
- Heuristic:OpenRLHF OpenRLHF Off Policy IS Correction Tip
- Heuristic:Huggingface Alignment handbook DPO Beta Selection
Environments
- Environment:Vllm project Vllm CUDA GPU Runtime
- Environment:Heibaiying BigData Notes Hadoop CDH Environment
- Environment:Dotnet Machinelearning Dotnet SDK And Runtime
- Environment:Turboderp org Exllamav2 Flash Attention Backend
- Environment:Allenai Open instruct Docker Container
- Environment:BerriAI Litellm Docker Deployment
- Environment:DataTalksClub Data engineering zoomcamp Dlt BigQuery Environment
- Environment:Microsoft Autogen LLM Provider API Keys
- Environment:Huggingface Diffusers Training Environment
- Environment:Norrrrrrr lyn WAInjectBench External Repos Dependencies