Posts by Ayush_SIngh

@Ayush_SIngh

Ayush Singh

Building open source infrastructure to detect LLM hallucinations, prompt attacks...
India github.com/AyushSingh110 Joined May 2026
1.9k Points50 Badges11 Connections11 Followers11 Following

Comments by Ayush_SIngh

2 days Articles 5 min read
If you have shipped an LLM agent, you have watched one go sideways in real time. Halfway through a task, it takes one bad step, and now it is confidently marching toward a wrong answer. The whole industry has gotten good at detecting that moment sp...
Jun 16 Articles 5 min read
AI agents rarely fail in a clean, obvious way. They do not always crash. They do not always throw an error. They do not always say, "I could not complete the task." Sometimes they fail more quietly. They give a confident answer with weak evidence....
post-cover-20606
Jun 12 Articles 3 min read
Nobody tells you what building alone actually feels like. The blog posts make it sound clean. You have an idea, you build it, you launch. Maybe you hit some technical walls, you push through, and eventually things work out. What they skip is the pa...
post-cover-20218
Jun 12 Articles 3 min read
Nobody tells you what building alone actually feels like. The blog posts make it sound clean. You have an idea, you build it, you launch. Maybe you hit some technical walls, you push through, and eventually things work out. What they skip is the pa...
post-cover-20218
Jun 6 Articles 1 min read
Most security systems are evaluated on attacks they have already seen. I decided to test mine on ones it hadn't. The Setup I built FIE "an open-source adversarial prompt detector for LLMs". 11 detection layers run in parallel on every incoming prompt...
Jun 4 Articles 3 min read
Most developers know they shouldn't use production data in non-production environments. But knowing and doing are two different things — and the gap between them is getting more expensive. According to Nick Mathison, Senior Product Manager for Delph...
post-cover-19698
May 15 Articles 2 min read
Most Dev's Journey stories start with a first line of code. Mine starts somewhere different. I came from marketing. And somewhere along the way, I became a technology writer — not because I could build software, but because I was genuinely curious a...
post-cover-17847
May 14 Articles 2 min read
Most Dev's Journey stories start with a first line of code. Mine starts somewhere different. I came from marketing. And somewhere along the way, I became a technology writer — not because I could build software, but because I was genuinely curious a...
post-cover-17847
May 12 Articles 4 min read
You built an AI feature. It works great in testing. Then someone types the wrong thing and your model does something it was never supposed to do. Here are the real attacks happening against LLMs right now, and how I built an open source system to c...
May 10 Articles 11 min read
LLMs are becoming part of real products now. They answer customers, summarize documents, write code, search internal knowledge bases, and make decisions inside workflows. But most LLM apps still have a quiet problem: > We usually find the failure a...
post-cover-17284
May 10 Articles 11 min read
LLMs are becoming part of real products now. They answer customers, summarize documents, write code, search internal knowledge bases, and make decisions inside workflows. But most LLM apps still have a quiet problem: > We usually find the failure a...
post-cover-17284
May 10 Articles 11 min read
LLMs are becoming part of real products now. They answer customers, summarize documents, write code, search internal knowledge bases, and make decisions inside workflows. But most LLM apps still have a quiet problem: > We usually find the failure a...
post-cover-17284
May 10 Articles 11 min read
LLMs are becoming part of real products now. They answer customers, summarize documents, write code, search internal knowledge bases, and make decisions inside workflows. But most LLM apps still have a quiet problem: > We usually find the failure a...
post-cover-17284
chevron_left