//about.md·whoami

# About

Headshot.jpg
Harikeshav Rameshkumar
$ whoami

Harikeshav Rameshkumar

@Harikeshav-R

Systems & AI/LLM Engineer

I build fast, low-level systems and production AI.

location
Columbus, OH

>a bit more about me

I'm a computer science student at The Ohio State University who likes the hard parts of software: distributed systems, compilers, cryptography, and shipping AI that actually holds up in production.

This past year I've cut LLM extraction latency by 74% on GE Aerospace's financial pipelines, hand-written AVX2/NEON kernels for a distributed inference engine that runs Llama 70B across consumer hardware, and built a fully-homomorphic-encryption ML runtime in Rust that keeps inputs encrypted end-to-end.

When I'm not shipping production code I'm usually at a hackathon — I've placed at RevolutionUC, TartanHacks, NextHacks, and HackOHI/O with teammates I trust, building everything from clinical-trial safety platforms to gamified finance twins.

education.yaml
school:The Ohio State University
degree:B.S. Computer Science
gpa:3.9 / 4.0
graduation:May 2028
location:Columbus, OH
honors:
  • Dean's List (all 4 semesters)
  • University Honors (all 4 semesters)
coursework:
Deep Learning & AISystems ProgrammingOperating SystemsComputer NetworkingData Structures & AlgorithmsFull Stack Web DevDatabase SystemsPrinciples of Programming LanguagesComputer OrganizationDiscrete StructuresLinear Algebra & Diff. Eq.Digital Logic
//experience/·git log --author

# Experience

  1. GE Aerospace

    Digital Technology Intern · AI / FinFlow Team

    Jun 2026 – Aug 2026Bengaluru, India

    Shipped production LLM contract-extraction pipelines on AWS Bedrock + Claude, financial-lineage AI systems on Neptune, and adoption analytics in Polars/DuckDB across GE Aerospace's financial platform.

    74%processing time cut
    726K+logs in 15s
    AWS BedrockClaudeAmazon NeptuneFastAPIPolarsDuckDBPython 3.12
  2. Wayfair

    AI Automation Apprenticeship

    Dec 2025 – Feb 2026Remote

    Built automated AI research and competitive intelligence agents in n8n and delivered a Market Intelligence Dashboard with automated pipelines for the Rugs team.

    70%research time cut
    50+products tracked
    n8nAI AgentsAutomated PipelinesMarket IntelligencePython
  3. Palantir Technologies

    Technical Fellow

    Nov 2025 – Jan 2026Remote

    Built end-to-end data workflows and interactive data applications in Palantir Foundry, creating ingestion pipelines, Ontology objects, and visualizations.

    50%faster data discovery
    10K+records processed
    Palantir FoundryOntologyData PipelinesObject TablesData Visualization
  4. Siage Solutions

    Software Engineering Intern

    Jun 2025 – Aug 2025Bengaluru, India

    Built enterprise RAG assistants, fine-tuned BERT triage engines, and high-throughput vLLM microservices on AWS EKS.

    92%search latency cut
    4.8×vLLM throughput
    LangChainQdrantOpenAIvLLMAWS EKSBERTPython
  5. Indian Institute of Technology, Madras

    Research Intern · Wireless Networks & Spatial AI Group

    Jun 2023 – Oct 2023Remote

    Researched physics-informed machine learning for indoor wireless signal propagation, predicting indoor Wi-Fi signal strength to replace manual site surveys.

    R²=0.975prediction accuracy
    86%less survey effort
    PythonTensorFlowScikit-LearnGaussian Process RegressionCNNLSTM
//projects/·ls -la

# Projects

// featured — 6 pinned

Penumbra-FHE

penumbra-fhe.rs

Encrypted ML inference engine (Rust + Python)

A privacy-preserving ML inference library that runs PyTorch / sklearn / XGBoost models under Fully Homomorphic Encryption — inputs and outputs never leave ciphertext. A 'narrow-waist' 3-layer architecture lowers ONNX graphs to a versioned IR executed on tfhe-rs primitives, with a PyO3 bridge and bit-for-bit exactness guarantees.

8 opshand-mapped to tfhe-rs
0-trustencrypted end-to-end
Rusttfhe-rsPyO3ONNXPyTorchRayonPython 3.12

LEAP

leap.cpp

Distributed LLM inference engine in C++20

A distributed LLM inference engine that runs models like Llama 3 70B across a ring of heterogeneous consumer devices using pipeline parallelism — bypassing the VRAM wall. Hand-written AVX2/NEON SIMD kernels, a zero-copy Linux kernel-module transport, and a custom INT8/FP32 weight serializer.

4×smaller footprint (INT8)
0-copykernel transport
C++20AVX2 / NEONOpenMPLinux kernel moduleLibTorchCMake

Understudy

understudy.py

Twin-fleet incident response agent with formal safety verification

An incident-response agent that rehearses candidate remediations across isolated Kubernetes twin environments receiving live-mirrored traffic before touching production. A deterministic rubric scores candidates in parallel, while a Z3-backed SMT safety kernel formally verifies system invariants to veto unsafe actions.

Z3 SMTformal safety veto
N-twinfleet rehearsal
Python 3.12KubernetesZ3 SMTLangGraphFastAPIPrometheusPostgreSQLDocker / k3d

Distill

distill.py

Intelligent LLM context-compression engine

3rd Place — The Token Company Track, NextHacks 2026

A high-performance LLM input-compression framework. A fine-tuned BERT token classifier scores every token by semantic necessity, then a two-tier pruning engine (context + token level) cuts prompt size while a zero-hallucination constraint engine preserves code and structure.

68%token reduction
~1.3%accuracy loss
PyTorchBERTHugging FacetiktokenFastAPIReact 19Chrome MV3

ArbOS

arb-os.rs

Low-latency prediction-market logical arbitrage engine

A low-latency algorithmic trading system that detects and exploits mathematical inconsistencies across 1,000 live Polymarket order books in ~14ms. An LLM reasoning core (watsonx.ai + LangGraph) maps logical relationships across prediction markets, coupled with a Rust execution engine featuring Fill-or-Kill settlement, circuit breakers, and fixed-point math.

~14msdetection latency
1,000live order books
RustTokioFastAPILangGraphPolygon PoSPostgreSQLReact 19watsonx.ai

Atlas

atlas.py

Local-first terminal-native AI job-application co-pilot

A local-first, terminal-native job-search platform featuring an interactive Textual TUI, Typer CLI, and background daemon over SQLite (WAL). A pluggable provider abstraction orchestrates Claude Code, Codex, and LiteLLM models with a 5-rung recovery ladder, truth-anchored resume tailoring, and a 100% line and branch test coverage gate.

100%branch test coverage
0-leaklocal-first privacy
Python 3.12Textual TUITyperSQLModelSQLite WALLiteLLMWeasyPrintPlaywright
//skills.toml·dependencies

# Skills

# skills.toml · 5 tables · 34 entries

[languages]8

  • Python
  • C++20
  • Rust
  • TypeScript
  • C
  • C#
  • Kotlin
  • SQL

[ai.llm]8

  • AWS Bedrock
  • Anthropic Claude
  • LangChain / LangGraph
  • PyTorch
  • RAG
  • vLLM
  • Hugging Face
  • pgvector / Qdrant

[systems]6

  • SIMD (AVX2/NEON)
  • Linux kernel modules
  • FHE (tfhe-rs)
  • OpenMP
  • Distributed systems
  • POSIX sockets

[cloud.infra]6

  • AWS (ECS/Lambda/Neptune)
  • Docker
  • Kubernetes
  • PostgreSQL
  • Redis
  • GitHub Actions

[backend.data]6

  • FastAPI
  • Polars
  • DuckDB
  • SQLAlchemy
  • Celery
  • Playwright
//awards.md·git tag --list

# Awards

$git tag --list7 awards4 podium finishes
2nd

2nd Place — Medpace Track

RevolutionUCUniversity of Cincinnati2026sponsor: Medpace

→ project: Pulse

60+ teams · 300+ participants

Recognized by Medpace clinical-research judges for a patient-safety ecosystem combining biometric anomaly detection with conversational symptom reporting and automated MedDRA/CTCAE grading.

3rd

3rd Place — Visa Track

TartanHacksCarnegie Mellon University2026sponsor: Visa

→ project: Penny

1,200+ hackers · 300+ projects

Praised by Visa technical leads for uniting multimodal receipt parsing, historical spending vector search, and a real-time 'true cost' e-commerce browser extension.

3rd

3rd Place — The Token Company Track

NextHacksCarnegie Mellon University2026sponsor: The Token Company

→ project: Distill

2,000+ participants · 400+ entries

Awarded for high-efficiency LLM infrastructure — a 68% token reduction and 37% latency drop while maintaining 99% accuracy across massive context windows.

Runner-Up

Best AI Hack Runner-Up

HackOHI/OThe Ohio State University2025

→ project: LeadForge

200+ teams · 800+ participants

Honored at OSU's flagship competition for agentic AI architecture — multi-agent lead scoring, autonomous React prototype generation, and real-time bi-directional AI voice calling.

1st

Game Development Track Winner

World Language AppathonOSU Department of Linguistics2025sponsor: Meta

→ project: VR Market Simulator

10 selective teams

Top game-dev honors from Meta and OSU Linguistics for immersive language learning — Meta Quest hand-tracking, dynamic VR budgeting, and NavMesh crowd AI.

Various

Multiple Podium Finishes

Regional HackathonsTISBHacks · NPSKRM Hacks · Regional Summits2022–2024

5+ events · 100–200 participants each

Repeated recognition for rapid full-stack execution and cryptographic security implementation under tight 24–48 hour competitive sprints.

Honor

Dean's List & University Honors

Academic DistinctionThe Ohio State University2023–present

All 4 completed semesters

Awarded Dean's List and University Honors every semester while maintaining a 3.9 / 4.0 cumulative GPA in Computer Science.

//blog/·ls -t | head -3

# Blog

//contact.md·./connect.sh

# Contact

>got a role, a project, or just want to say hi? my inbox is open.

bash — connect.sh

$whoami

Harikeshav Rameshkumar — Systems & AI/LLM Engineer

// open to SWE / AI internship + new-grad conversations

// built from scratch with react + vite + tailwind, styled after neovim + catppuccin

$ git commit -m "shipped" · harikeshav.me