Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Mage ai Mage ai Building a New Destination Connector
- Workflow:Apache Paimon Lance Format Analytics
- Workflow:Guardrails ai Guardrails Server Deployment
- Workflow:Huggingface Alignment handbook ORPO Single Stage Alignment
- Workflow:Nightwatchjs Nightwatch Component Testing
- Workflow:Microsoft Playwright End to end test authoring
- Workflow:Openai Openai node Streaming To Client
- Workflow:Tensorflow Serving Model Version Management
- Workflow:PeterL1n BackgroundMattingV2 Video matting inference
- Workflow:Pola rs Polars Lazy Query Pipeline
Principles
- Principle:Predibase Lorax Structured Response Parsing
- Principle:Promptfoo Promptfoo Telemetry Collection
- Principle:AUTOMATIC1111 Stable diffusion webui Sampling Architecture
- Principle:Explodinggradients Ragas Prompt Iteration Comparison
- Principle:Tensorflow Serving TFRT Model Management
- Principle:Deepspeedai DeepSpeed Inference Engine Init
- Principle:Datajuicer Data juicer Configuration Initialization
- Principle:Webdriverio Webdriverio Global API
- Principle:Apache Kafka Release Tag And Vote
- Principle:ARISE Initiative Robosuite Camera Randomization
Implementations
- Implementation:OpenRLHF OpenRLHF ValueLoss
- Implementation:Eventual Inc Daft DataFrame Groupby
- Implementation:Online ml River Imblearn HardSampling
- Implementation:Protectai Llm guard Output BanCode
- Implementation:SeleniumHQ Selenium Closure String
- Implementation:Vllm project Vllm Broadcast Load Epilogue Array C3X
- Implementation:ClickHouse ClickHouse Data Lakes Importer
- Implementation:ArroyoSystems Arroyo Avro Schema Converter
- Implementation:Apache Shardingsphere ShowProcessListHandler Handle
- Implementation:CarperAI Trlx GPTRewardModel
Heuristics
- Heuristic:SqueezeAILab ETS Embedding Model GPU Collocation
- Heuristic:Mlc ai Mlc llm OpenCL Memory Floor Workaround
- Heuristic:Eric mitchell Direct preference optimization FSDP Mixed Precision BFloat16
- Heuristic:Unstructured IO Unstructured Golden File Diff
- Heuristic:Apache Kafka Coordinator Loading Commit Interval
- Heuristic:Spcl Graph of thoughts Backoff Retry On API Errors
- Heuristic:PeterL1n BackgroundMattingV2 Training Batch Size And Resolution
- Heuristic:Huggingface Transformers Label Smoothing Multi Label Warning
- Heuristic:Arize ai Phoenix AIMD Concurrency Control
- Heuristic:OWASP Www project top 10 for large language model applications SHA Pinning For GitHub Actions
Environments
- Environment:Iamhankai Forest of Thought Python CUDA Runtime
- Environment:Ggml org Ggml C Cpp Build Environment
- Environment:Mlflow Mlflow MLflow Server Environment
- Environment:Run llama Llama index Sentence Transformers Finetuning
- Environment:Sgl project Sglang CUDA Runtime
- Environment:Intel Ipex llm Portable Environment
- Environment:OpenRLHF OpenRLHF Ray Distributed Environment
- Environment:Huggingface Transformers PyTorch 24 CUDA
- Environment:Wandb Weave Python SDK Runtime
- Environment:Mlfoundations Open flamingo Evaluation Dependencies