Pinned Loading
-
rlhf-from-scratch-on-distilgpt2
rlhf-from-scratch-on-distilgpt2 PublicRLHF pipeline from first principles: SFT, LoRA, reward modeling, PPO, and DPO/IPO/KTO/ORPO/SimPO, unit tested and run end to end on DistilGPT2
Python
-
llm-agent-calibration-eval-harness
llm-agent-calibration-eval-harness PublicLLM eval / RL-environment design harness: calibrating agent failure modes across 5 forecasting tasks
-
scaling-llms-distributed-systems
scaling-llms-distributed-systems PublicStudy notes on scaling LLM training, distributed systems, and inference, worked through the JAX scaling book end to end.
-
German-Product-Title-NER
German-Product-Title-NER PublicProduction-grade multilingual NER system for structured attribute extraction from German e-commerce titles using XLM-R + CRF with modular training, HPO, and post-processing pipelines.
Python
-
Kenda_Tires_Sales_Forecasting
Kenda_Tires_Sales_Forecasting PublicEnd to end SKU level sales forecasting system built for a tire distributor. LightGBM global model, Croston SBA for intermittent demand, MinT hierarchical reconciliation, and A to D confidence gradi…
Jupyter Notebook
-
Shell-Machine-Learning-Fuel-Blend-Prediction
Shell-Machine-Learning-Fuel-Blend-Prediction PublicDeveloped an end-to-end ML pipeline for fuel blend property prediction using advanced preprocessing, blend-weighted features, nonlinear transformations, PCA, and clustering. Built a LightGBM + Ridg…
Jupyter Notebook
If the problem persists, check the GitHub status page or contact support.