{{ heroTitleEls }}
through {{ heroWordEl }}
I use engineering, research and practical experimentation to build AI systems that are useful, inspectable and easier to trust.
{{ p.title }}
{{ p.blurb }}
{{ p.title }}
{{ p.blurb }}
Other
Projects
{{ tableHintText }}
What I bring to the table
Learn more about me!
Complex systems are my playground, useful software is my goal.
I'm an AI systems engineer, AI agent developer and founder of Nexasity AI focused on turning complicated workflows into useful software. I build AI agents, RAG systems, voice workflows and evaluation tools, always trying to make the final system clearer, more reliable and easier to use.
Playground
Experiments, side quests and things I made to understand how a system works.
Projects
A few systems, experiments and research projects I have built.
{{ p.title }}
{{ p.blurb }}
{{ p.title }}
{{ p.blurb }}
From messy workflows to reliable systems
I started by building software, websites and small experiments, then kept moving closer to the problems where AI could actually be useful.
The
Now I build AI agents, RAG systems, voice workflows and software tools through Nexasity AI.
Outside of work, I’m almost always building something.
I like projects with a little >/{{ termText }}|. A good system should do something useful, but it should also make its decisions easier to inspect, test and improve. Although...
A few things I have been investigating
{{ favCategoryTitle }}
False Containment: Measuring the Evidence Required to Verify Autonomous AI Incident Resolution
To be verified · 2026 · Apart Research AI Incident Response Sprint
A research project about how much evidence is actually needed before claiming that an autonomous AI incident has been resolved.
When the Record and the Report Diverge: Self-Report Fidelity Collapses Under Structured Provenance in Claude Haiku 4.5
Daud Ibrahim Hassan, Deniz Chen, and Soumya Parthasarathy · 2026 · Apart Research
A study of whether language models accurately report what they did when structured provenance and tool-use records are available.
MarkLens: Measuring Cross-Manager Valuation Dispersion in Private Credit at Scale ↗
Daud Ibrahim Hassan · 2026 · Independent preprint
A reproducible SEC EDGAR study of cross-manager private-credit valuation differences at scale.
Monitor Calibration Transport: Evaluating AI Safety Monitor Reliability Across Models and Distribution Shifts
Daud Ibrahim Hassan · 2026 · Independent AI safety research
A study of whether safety monitors calibrated on one model or task remain reliable when the model, task or deployment conditions change.
The Secret Life of Walter Mitty
Ben Stiller
Mr. Robot
Sam Esmail
Palm Springs
Max Barbakow
Pantheon
Craig Silverstein
{{ caseTitle }}
{{ caseSubtitle }}
Role
{{ caseRole }}
Impact
{{ caseImpact }}
Overview
{{ caseOverview }}
Problem
{{ caseProblemIntro }}
- {{ pt }}
The Opportunity
{{ caseOpportunity }}
Scope & Reach
{{ sc.label }}
{{ sc.value }}
What I owned
- {{ o }}
Color System
{{ col.label }}
{{ it.desc }}
{{ col.text }}
Objectives → Design framing
{{ ob.question }}
Workflow Efficiency
{{ line }}
Designers could now…
{{ line }}
Key constraints & tradeoffs
Research inputs
- {{ r }}
{{ caseObservationsHeading }}
{{ caseObservationsLeftLabel }}
{{ blk.heading }}
- {{ line }}
{{ blk.text }}
{{ caseObservationsRightLabel }}
{{ blk.heading }}
- {{ line }}
{{ blk.text }}
How objectives translated to features
Why?
{{ caseSolutionActive.copy }}
{{ caseSolutionActive.listLabel }}
- {{ tr }}
{{ caseSolutionActive.paragraphLabel }}
{{ caseSolutionActive.thinking }}
KPIs & measurement strategy
{{ k.label }}
Why this mattered
{{ k.why }}
What was measured
- {{ w }}
Why this was chosen
{{ k.reason }}
Solution
{{ sec.section }}
{{ it.copy }}
Outcome & Reflection
{{ grp.label }}
Results
- {{ line }}
Impact
{{ grp.impact }}
If I had more time…
{{ m.title }}
{{ m.copy }}
{{ caseReflection }}
{{ para }}
This project highlights my strengths in:
- {{ s }}
Have a system in mind?
I’d love to hear what you are trying to build.
Dhaka, Bangladesh


