Skip to content
View akashm776's full-sized avatar

Block or report akashm776

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
akashm776/README.md

Hi, I’m Akash Mittal

I work on research-shaped AI engineering problems in optimal transport, multimodal learning, and reliable AI systems.

My strength is naive, stubborn questioning. I question assumptions even when they are stated confidently or treated as given. I like translating technical concepts into language I can actually reason with, then testing whether the idea still survives in code.

My thesis work focused on warm-start algorithms for bipartite matching and optimal transport. Here are the projects I'm currently working on in my personal time:

State-Conditioned Synthetic Supervision (SCSS), formerly OTCO, asks: can we store reusable functions that generate useful training signals, rather than relying only on stored examples? The key is the learner’s state: what helps one model at one point in learning may hurt it at another. I started with OT-based synthetic hard negatives; the current experiments use matched CLIP continuations to separate geometric hardness, immediate update effects, and sustained learning outcomes. The broader goal is to learn when and how to use generated supervision—not to assume that more synthetic data is better. Reduced data requirements and reliable downstream gains remain open questions. Research program · Experiments and results.

Exactness-Triage comes from a similar instinct. Agent memory and context compaction often lean on summarization because LLMs are good at it. I wanted to test the assumption behind that move: some tool outputs may contain exact, load-bearing details (a key name, a failing diff, a specific error string) that should be preserved before compression, not reconstructed after summarization.

The common thread: I like questioning the default interpretation of a technique, then building the smallest experiment I can to see whether the idea survives contact with reality.

You can reach me at akashmit28@gmail.com.

Pinned Loading

  1. SCSS SCSS Public

    State-Conditioned Synthetic Supervision: research on reusable generators of training signals and their state-dependent utility, with matched CLIP experiments and reproducible evidence.

    Python

  2. exactness-triage exactness-triage Public

    Controlled experiments on preserving exact, load-bearing tool outputs while compressing other observations in AI coding-agent context.

    Python

  3. workflow-harness workflow-harness Public

    Deterministic workflow-governance prototype with policy validation, scoped approvals, and auditable artifacts; V1 is safe no-op only, with no real tool execution.

    Python

  4. warm-start-ot warm-start-ot Public

    In-progress reference implementation of thesis algorithms for warm-start bipartite matching and optimal transport, with Java solver cores and planned Python bindings.

    Java

  5. Divide-and-Conquer-Approach-to-OT Divide-and-Conquer-Approach-to-OT Public

    Java implementations and experiments comparing traditional and divide-and-conquer Hungarian search for minimum-cost bipartite matching and optimal transport.

    Java