Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Cypress io Cypress E2E Test Execution
- Workflow:Ggml org Llama cpp Model Perplexity Evaluation
- Workflow:Apache Flink Async Sink Lifecycle
- Workflow:Diagram of thought Diagram of thought DoT Trace Extraction
- Workflow:Avhz RustQuant Model Calibration
- Workflow:Huggingface Open r1 Model Evaluation
- Workflow:Eric mitchell Direct preference optimization Custom Dataset Integration
- Workflow:DevExpress Testcafe CLI Test Execution
- Workflow:SeleniumHQ Selenium Selenium Grid Deployment
- Workflow:Facebookresearch Habitat lab Agent Benchmarking
Principles
- Principle:Microsoft Autogen Team Execution
- Principle:LLMBook zh LLMBook zh github io Preference Data Preparation
- Principle:Apache Dolphinscheduler Datasource Parameter Configuration
- Principle:Scikit learn contrib Imbalanced learn Random Under Sampling Boosting
- Principle:Apache Paimon Lazy Blob Loading
- Principle:Neuml Txtai Content Conversion
- Principle:Sgl project Sglang OpenAI Client Configuration
- Principle:Microsoft Semantic kernel Subprocess Composition
- Principle:Microsoft DeepSpeedExamples Baseline PyTorch Training
- Principle:MaterializeInc Materialize Composition Service Definition
Implementations
- Implementation:FMInference FlexLLMGen Policy
- Implementation:EvolvingLMMs Lab Lmms eval Alpaca Audio Utils
- Implementation:Apache Hudi OptionsResolver Write Configuration
- Implementation:Apache Kafka Docker Buildx Remove
- Implementation:EvolvingLMMs Lab Lmms eval ChatMessages
- Implementation:NVIDIA DALI EfficientDet Preprocessor
- Implementation:Tensorflow Serving Tfrt Multi Inference
- Implementation:Online ml River Tree HoeffdingTree
- Implementation:Online ml River Stats Minimum
- Implementation:Openai Openai python Vector Store Create Params
Heuristics
- Heuristic:Huggingface Alignment handbook Global Batch Size Scaling
- Heuristic:Dagster io Dagster Mandatory Ruff Formatting
- Heuristic:Interpretml Interpret EBM Hyperparameter Tuning Guide
- Heuristic:Guardrails ai Guardrails Guard History Memory Management
- Heuristic:Eric mitchell Direct preference optimization TF32 Matmul Precision
- Heuristic:BerriAI Litellm SSL Cipher Optimization
- Heuristic:Sail sg LongSpec NCCL Distributed Settings
- Heuristic:Onnx Onnx Big Endian Byte Order Handling
- Heuristic:Cypress io Cypress Timeout Tuning
- Heuristic:Sdv dev SDV HMA Schema Simplification
Environments
- Environment:Apache Shardingsphere Java Runtime Environment
- Environment:Puppeteer Puppeteer Cross Platform Browser Environment
- Environment:EvolvingLMMs Lab Lmms eval GPU Compute Environment
- Environment:LMCache LMCache NIXL Transfer Library
- Environment:Unstructured IO Unstructured Ingest CLI
- Environment:Microsoft Onnxruntime Sklearn Conversion Environment
- Environment:Mistralai Client python Python SDK Environment
- Environment:Duckdb Duckdb Code Generation Tools
- Environment:Apache Airflow Kubernetes Helm Environment
- Environment:VainF Torch Pruning CUDA GPU Benchmarking