Anmol Gorakshakar

Computational Biology / Platform Engineering

Scientific systems that survive production.

Current Focus

Biology, ML, and infrastructure in the same room.

The work sits between scientific reasoning and production reliability: molecular modelling, NLP, data contracts, APIs, relational schemas, and operational pipelines.

Featured Projects

Selected work

Full timeline

VSE

Core Platform

Built the Kubernetes-deployed scoring platform that delivers every microbiome health score to customers and clinicians, replacing VM-era scoring with auditable release infrastructure.

ProductionAKSKEDATerraformScoring

DASH

Internal Leverage

Standardized score release workflows behind compiled Python CLIs, reducing operational friction across deploy, rollback, notification, export, and ID translation tasks.

Published internallyCLIRelease workflowNuitka

HR

Partner Platform

Rebuilt the Historeceptomics-Algorithm as a Python package and relational database, making it over 1000x faster and usable through Molsoft ICM partner tooling.

Live1000xMolsoft ICMAbbVie
ChatGPTCompatible across the spectrum

Scoring Copilot

Proposes a mechanism. Then tries to prove itself wrong.

A real run of the research pipeline that discovers and grades candidate scoring features — one agent proposes a mechanism-of-action hypothesis, a second tries to falsify it, and the survivors are graded across five evidence axes before they reach production scoring.

scoring-copilot · feature discovery

sql-executor

Named, parameterized, pre-reviewed SQL queries for common lookups, plus a raw-SQL escape hatch for one-off schema exploration.

translate-ids-mcp

Cross-translates identifiers (kit-id, user-id, test-id, external-id) across internal databases, optionally joined to analysis-id.

expression-mcp

Pulls filtered expression data (by feature type and read-count threshold) from the bioinformatics database as a clean, ready-to-analyze CSV.

cohort-exploration-mcp

Queries Viome's internal biological and customer metadata to check whether a hypothesized relationship actually shows up in real cohort and label data.

kegg-mcp

Looks up pathway and mechanism-of-action data from KEGG — the biological grounding for a hypothesis, not just a citation.

literature-review-mcp

Full-text literature retrieval with custom embeddings and ranking — built because abstract-only search missed too much mechanism-level nuance.

GitHub Activity

Past year

GitHub
767 Contributions
72 Pull Requests
40 Repos

Language Mix

  • Python43.0%
  • Java38.3%
  • HCL9.4%
  • Shell2.8%
  • R2.0%
  • Go1.6%
  • Other2.9%

What I Build

Production-grade scientific platforms.

Flask APIs, relational databases, scalable curation scripts, NLP models, multiomics analysis workflows, and molecular simulation tooling with clear handoff boundaries.

Looking For

Senior data science or platform roles.

Best fit: biotech, health-tech, AI drug discovery, precision medicine, or infrastructure teams building for scientific workloads.