P/01
proprietary
GenAI Medical Summarisation
A production RAG pipeline that reads dense clinical documentation and returns the parts
a neurologist actually needs. Semantic chunking, retrieval and prompt orchestration
tuned for clinical precision — extraction that used to take minutes takes seconds.
LangChainorchestration
Semanticchunking
Clinicaldomain
RAGLLMsLangChainHealthcare
P/02
proprietary
LLaMA Rummy Bot
A LLaMA-driven opponent for Junglee's Learn & Practice mode — it plays a credible
hand and explains why, turning onboarding from a tutorial into a game.
LLaMAGenAIGame Systems
P/03
proprietary
A/B & Recommendation MLOps
MLOps pipelines at Junglee Games for A/B testing and dynamic recommendation
systems — models fed from the same governed tables that back the reporting.
MLOpsA/B TestingRecommenders
P/04
proprietary
Pharma MDM Pipeline
Entity resolution at scale: 50M+ records collapsed into Golden Records with fuzzy
matching and NLP, hitting 87% match accuracy on notoriously dirty HCP/HCO data.
MDMSparkNLPAxtria
data-dbt
A production-grade dbt project for e-commerce analytics — 11 models, 56 tests, CI/CD,
BI integration and an interactive GitHub Pages site. Built on dbt + DuckDB, so the
whole warehouse runs on a laptop.
11models
56tests
CI/CDevery PR
dbtDuckDBAnalytics Engineering
data-builder
A visual ETL platform — connect databases, browse catalogs, build pipelines by drag and
drop, stream CDC to S3, then schedule, monitor and export logs. FastAPI + React.
ETLCDCFastAPIReact
datalearn
LeetCode-style SQL practice that runs entirely in the browser — every solution
validated by DuckDB-WASM, plus a learning hub and admin CMS. Next.js, Prisma, Monaco.
DuckDB-WASMNext.jsSQLEducation
learn_llm_code
Build an LLM from scratch: 15 runnable Python scripts that walk from a bare neural
network to a 2026-architecture, DPO-aligned tiny GPT — with a companion site.
LLMDPOPyTorchFrom Scratch
pytq
TurboQuant for PyTorch — near-optimal vector quantization for LLM KV-cache compression.
The unglamorous inference work that decides whether a model is affordable to serve.
QuantizationKV CachePyTorchInference
r-llm
Python vs Rust, honestly measured: an OpenAI-compatible batch-summariser CLI built
twice, with a deterministic mock server, cross-language golden checks and benchmarks.
RustPythonBenchmarksLLM Ops
open-desktop-gpt
Your LLM compiles the wiki, you read it. A desktop app inspired by Karpathy's
LLM-knowledge-base idea — generation as a reading surface, not a chat log.
LLMTypeScriptDesktop
papers-code
Implementations of Standard ES and EGGROLL from Evolution Strategies at the
Hyperscale (arXiv:2511.16652). Reading a paper properly means running it.
ResearchEvolution StrategiesPython
Wikipedia Citation Verifier
Retrieval before it was fashionable — a Chrome extension that embeds a cited document
and surfaces the sentence closest to the claim, so you can check whether the citation
says what it claims. IIIT Delhi IR project.
Embeddingsdoc2vecChrome ExtIR
easy-pdf
Free online PDF tools — merge, split, compress, convert, OCR. 100% client-side, so
files never leave the device. The privacy guarantee is architectural, not a promise.
Client-sidePDFJavaScript
nothing here yet — try another filter.