Skip to content
View zahid23saim's full-sized avatar

Block or report zahid23saim

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
zahid23saim/README.md

Hi, I'm Zahid Ahmed 👋

Software Developer · AI Evaluation & Data Quality — based in Maharashtra, India.

Full-stack developer with an MCA and 3+ years of dedicated Python development, now focused on large-language-model evaluation — designing tasks and tests that reveal where AI coding agents fail. I build in Python, .NET, and SQL, and I write about it.

🔗 Portfolio · ✍️ Technical Writing · 💼 LinkedIn · ✉️ zahid23saim@gmail.com


🛠️ Tech

Python C# .NET Java SQL Oracle JavaScript Docker

Plus: LLM evaluation & benchmark authoring, factuality & STEM auditing, test writing, database optimization.


📌 Featured Projects

Project What it is Built with
llm-eval-harness A dependency-free harness for scoring LLM answers against a gold set — per-question match rules, validation, CI-friendly. Live demo → Python · pytest
aspnet-minimal-api A tested ASP.NET Core (.NET 8) task API — full CRUD, validation, in-memory integration tests (7/7 passing). C# · ASP.NET Core · xUnit
plsql-bulk-examples Turning a slow row-by-row Oracle load into fast batched BULK COLLECT + FORALL, with SAVE EXCEPTIONS. Oracle PL/SQL

✍️ Writing


🌐 English · Hindi · Tamil · Marathi  |  Open to remote, project-based software & AI-evaluation work.

Pinned Loading

  1. aspnet-minimal-api aspnet-minimal-api Public

    A small, tested ASP.NET Core Minimal API (.NET 8): task-tracker CRUD with validation, correct status codes, and in-memory integration tests.

    C#

  2. llm-eval-harness llm-eval-harness Public

    A tiny, dependency-free Python harness for scoring LLM answers against a gold set (exact / contains / numeric matching, CI-friendly).

    Python

  3. plsql-bulk-examples plsql-bulk-examples Public

    Runnable Oracle PL/SQL examples: turning a slow row-by-row load into fast batched BULK COLLECT + FORALL, with SAVE EXCEPTIONS error handling.

    PLSQL

  4. zahid23saim.github.io zahid23saim.github.io Public

    Personal portfolio — Zahid Ahmed, software developer & AI evaluation specialist.

    HTML