Skip to content
View Josepheeeee's full-sized avatar

Block or report Josepheeeee

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Josepheeeee/README.md
Atlas — Systems Engineer

English · 简体中文

Explore projects GitHub

👋 Hi, I’m Atlas

AI/ML Systems Engineer · Agent & LLM

I build machine-learning systems that connect models with real tools, data contracts, evaluation loops, and reproducible engineering workflows.

What I work on

Agent systems LLM post-training NL2SQL
Document AI Data pipelines Evaluation

What I care about

Clear interfaces · deterministic guards
Leakage-safe evaluation · honest conclusions

✨ Selected work

LoRA SFT + veRL GRPO for long-horizon shopping agents that must search, verify, select variants, and purchase.

Highlight: Baseline 1.0% → SFT 57.0%; GRPO 58.5%, with the current SFT→GRPO gain not statistically significant.

Stateful natural-language querying with table routing, schema-aware semantic parsing, constrained DSL, deterministic SQL compilation, and read-only MCP tools.

Highlight: LLM handles semantic mapping; code owns SQL structure and safety boundaries.

Conservative financial-table structure repair using directional morphology, line detection, chunk-aware processing, and regression checks.

Highlight: Team result: B-rank #7; my focus was table-header repair and regression coverage.

A local, synthetic-data reproduction of a manufacturing quality-monitoring workflow from source tables to dashboard, alert outbox, and traceable results.

Highlight: Runnable backend + frontend demo with deterministic failure injection and recovery paths.

🧰 Toolbox

Python PyTorch LangGraph OpenCV SQL MCP React Git

🧭 How I build

question → explicit data contract → small reproducible run
        → deterministic checks → paired evaluation → honest conclusion

I prefer a useful limitation over an impressive but unsupported claim. Each project README includes the current implementation boundary and the fastest path for a technical reader to reproduce or inspect it.

Thanks for stopping by · Explore the repositories or start a conversation on GitHub.

Popular repositories Loading

  1. Learning_Projects Learning_Projects Public

    a place to store everything i learned on my way to be an algorithm engineer

    Python

  2. march-machine-learning-mania-2026 march-machine-learning-mania-2026 Public

  3. minimind-v-k12 minimind-v-k12 Public

    Python

  4. tiger-recommender tiger-recommender Public

    Reproducible semantic-ID recommendation baseline with deterministic evaluation.

    Python

  5. lookalike-ranking lookalike-ranking Public

    Leakage-safe audience expansion and ranking baseline.

    Python

  6. Josepheeeee Josepheeeee Public

    Personal GitHub profile README.