Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:HKUDS AI Trader Agent Decision Loop
- Workflow:Microsoft Onnxruntime Nodejs Inference
- Workflow:Huggingface Open r1 Dataset Pass Rate Filtering
- Workflow:Lucidrains X transformers Encoder Decoder Sequence to Sequence
- Workflow:Helicone Helicone Integrate Provider To Gateway
- Workflow:Webdriverio Webdriverio WDIO Testrunner Setup
- Workflow:Mistralai Client python Finetuning Job Management
- Workflow:Zai org CogVideo Diffusers Image to Video Inference
- Workflow:DataExpert io Data engineer handbook PySpark Iceberg Job Execution
- Workflow:Testtimescaling Testtimescaling github io Automated Citation Tracking
Principles
- Principle:Microsoft Agent framework Declarative Tool Binding
- Principle:Zai org CogVideo Text to Video Generation
- Principle:ARISE Initiative Robomimic Checkpoint Loading
- Principle:Ggml org Llama cpp Speculation Initialization
- Principle:Norrrrrrr lyn WAInjectBench JSONL Results Serialization
- Principle:Risingwavelabs Risingwave CDC Schema History
- Principle:Run llama Llama index Training Data Validation
- Principle:Spotify Luigi Pipeline Parameterization
- Principle:Microsoft Playwright Element Location and Interaction
- Principle:TA Lib Ta lib python C Library Installation
Implementations
- Implementation:Microsoft DeepSpeedExamples Vision Transformer Model
- Implementation:Astronomer Astronomer cosmos Aws Eks Operators
- Implementation:Trailofbits Fickling Find File Properties
- Implementation:Huggingface Transformers Time Generate Warmup
- Implementation:Microsoft Semantic kernel ImportPluginFromOpenApiAsync
- Implementation:Apache Dolphinscheduler BaseAdHocAndPooledClient Extension
- Implementation:Online ml River Bandit Envs KArmedTestbed
- Implementation:Recommenders team Recommenders Affinity Matrix
- Implementation:NVIDIA NeMo Curator ParquetWriter
- Implementation:Norrrrrrr lyn WAInjectBench text ensemble detect
Heuristics
- Heuristic:DataExpert io Data engineer handbook Flink Checkpointing Interval Tuning
- Heuristic:Sail sg LongSpec Triton Block Size Tuning
- Heuristic:Mlc ai Web llm Penalty Parameter Defaults
- Heuristic:Tensorflow Serving Batching Thread Tuning
- Heuristic:Lakeraai Pint benchmark Chunking Stride 25 Percent Overlap
- Heuristic:Langfuse Langfuse LLM Rate Limit 24h Abandon
- Heuristic:Tensorflow Serving Model Warmup Strategy
- Heuristic:Lance format Lance BM25 FTS Configuration
- Heuristic:Wandb Weave Concurrency Deadlock Prevention
- Heuristic:Intel Ipex llm QLoRA Training Hyperparameters
Environments
- Environment:Fastai Fastbook Sklearn Environment
- Environment:DevExpress Testcafe Chrome Browser
- Environment:Mit han lab Llm awq Python Runtime Environment
- Environment:Sktime Pytorch forecasting Matplotlib Plotting Dependencies
- Environment:AUTOMATIC1111 Stable diffusion webui Xformers Attention
- Environment:Dotnet Machinelearning OneDal Acceleration
- Environment:Protectai Llm guard API Server Deployment
- Environment:OpenRLHF OpenRLHF vLLM Environment
- Environment:Eventual Inc Daft Ray Distributed Runner
- Environment:Lucidrains X transformers Python Environment