Ask better questions.
Compare models using structured rubrics for accuracy, citations, and reasoning.
I build AI systems and investigate how well their answers hold up to evidence.
Contact meI’m an AI developer with a background in computer science and information systems. I work on language-model evaluation, retrieval-augmented generation, and practical tools for inspecting AI outputs.
See selected projectsCompare models using structured rubrics for accuracy, citations, and reasoning.
Connect generated claims to supporting passages in technical documentation.
Turn model outputs into reviewable tools for faculty and domain-specific knowledge work.
Across my projects, the work connects source material, retrieval, model outputs, and evaluation. The goal is to make errors visible and give people the context to judge an answer.
A closer look at the systems I’ve built and the questions behind them.
Compared language models on engineering questions using structured rubrics for accuracy, grounding, citation quality, and reasoning.
Python · Model APIs · Streamlit · Power BI
Created a labeled question dataset spanning hardware, networking, cloud, embedded systems, and cybersecurity. Built interactive views to compare models and review failure patterns.
A retrieval and verification pipeline that checks model-generated claims against engineering documentation.
LangChain · FAISS · Azure Cognitive Search · Azure Functions
Extracted factual assertions and retrieved supporting passages. Classified each claim as supported, unsupported, or contradicted, retaining evidence and confidence information for review.
Compared vision models on circuit diagrams, network topologies, and cloud architecture diagrams.
GPT-4o · Claude Vision · Azure AI Vision · JSON Schema
Structured model interpretations as JSON and compared them against ground-truth annotations. Reviewed component recall, description accuracy, and recurring interpretation errors.
A Claude-powered tool for university faculty to review individual contributions to group projects.
Python · FastAPI · Anthropic API · Azure SQL · Power BI
Combined activity logs and peer feedback into contribution summaries with reasoning traces. Iterated prompts around classroom edge cases and documented the system for faculty use.
A RAG assistant for crop-management questions, grounded in agricultural extension documents.
Python · LangChain · FAISS · Streamlit · Azure App Service
Added an evaluation harness for answer grounding, hallucinations, and task completion. Refined prompt templates through structured review as part of a computer science capstone.
Murray State University · 2025–2026
Lovely Professional University · 2020–2024
Tell me what you have in mind.
This form uses FormSubmit to forward your name, email, purpose and optional phone number to my inbox. FormSubmit retains submissions for 30 days and uses Google reCAPTCHA. Privacy and terms. You can also email me at deemag.work@outlook.com.