Nithin Mohan TK

Guardrails and Safety for LLMs: Building Secure AI Applications with Input Validation and Output Filtering

Posted on August 26, 2024 by Nithin Mohan TK 12 min read

Introduction: Production LLM applications need guardrails to ensure safe, appropriate outputs. Without proper safeguards, models can generate harmful content, leak sensitive information, or produce responses that violate business policies. Guardrails provide defense-in-depth: input validation catches problematic requests before they reach the model, output filtering ensures responses meet safety standards, and content moderation prevents harmful generations. […]

Read more →

Rate Limiting for LLM APIs: Token Buckets, Queues, and Adaptive Throttling

Posted on August 22, 2024 by Nithin Mohan TK 13 min read

Introduction: LLM APIs have strict rate limits—requests per minute, tokens per minute, and concurrent request limits. Exceeding these limits results in 429 errors that can cascade through your application. Effective rate limiting on your side prevents hitting API limits, provides fair access across users, and enables graceful degradation under load. This guide covers practical rate […]

Read more →

Achieving DevOps Harmony: Building and Deploying .NET Applications with AWS Services

Posted on August 20, 2024 by Nithin Mohan TK 5 min read

The Evolution of .NET Deployment on AWS After two decades of building enterprise applications, I’ve witnessed the transformation of deployment practices from manual FTP uploads to sophisticated CI/CD pipelines. When AWS introduced their native DevOps toolchain, it fundamentally changed how we approach .NET application delivery. The integration between CodeCommit, CodeBuild, CodePipeline, and ECR creates a […]

Read more →

LLM Security: Understanding Prompt Injection, Jailbreaking, and Attack Vectors (Part 1 of 2)

Posted on August 20, 2024 by Nithin Mohan TK 14 min read

A comprehensive guide to securing LLM applications against prompt injection, jailbreaking, and data exfiltration attacks. Includes production-ready defense implementations.

Read more →

LLM Batch Processing: Scaling AI Workloads from Hundreds to Millions

Posted on August 18, 2024 by Nithin Mohan TK 9 min read

Introduction: Processing thousands or millions of items through LLMs requires different patterns than single-request applications. Naive sequential processing is too slow, while uncontrolled parallelism hits rate limits and wastes money on retries. This guide covers production batch processing patterns: chunking strategies, parallel execution with rate limiting, progress tracking, checkpoint/resume for long jobs, cost estimation, and […]

Read more →

The Future of Work: How AI and Automation Are Reshaping Careers

Posted on August 17, 2024 by Nithin Mohan TK 8 min read

After two decades of architecting enterprise systems and leading digital transformation initiatives across financial services, healthcare, and technology sectors, I’ve witnessed firsthand how AI and automation are fundamentally reshaping the nature of work. This isn’t merely about replacing tasks—it’s about reimagining entire value chains, creating new categories of roles, and demanding a fundamental shift in […]

Read more →

Searching in

Author: Nithin Mohan TK

Guardrails and Safety for LLMs: Building Secure AI Applications with Input Validation and Output Filtering

Rate Limiting for LLM APIs: Token Buckets, Queues, and Adaptive Throttling

Achieving DevOps Harmony: Building and Deploying .NET Applications with AWS Services

LLM Security: Understanding Prompt Injection, Jailbreaking, and Attack Vectors (Part 1 of 2)

LLM Batch Processing: Scaling AI Workloads from Hundreds to Millions

The Future of Work: How AI and Automation Are Reshaping Careers