Lugman Hussain Khan

Co-Founder & Head of Engineering

LinkedInGithub
L

Lugman Hussain Khan

engineering
Four architecture approaches for using smaller language models

Explore four architectural approaches for using SLMs to reduce inference cost and latency while preserving application quality and reliability.

L

Lugman Hussain Khan

engineering
Improving agentic web search in small language models

Enhancing multi-hop web search in 2-billion-parameter language models via 193 distilled trajectories.

L

Lugman Hussain Khan

engineering
A practical guide to LLM inference optimization

A practical guide to LLM inference optimization, covering concurrency, context length, quantization, GPU memory, vLLM vs. SGLang, and latency-throughput trade-offs.

L

Lugman Hussain Khan

engineering
Move beyond traditional agentic search with specialised retrieval sub-agents

Traditional agentic search bloats LLM context windows. Discover how specialised retrieval sub-agents optimise multi-hop search for long-horizon AI workloads.

L

Lugman Hussain Khan

product
Building ElectionGPT for the 2026 Tamil Nadu elections

How Madhi AI built ElectionGPT, an AI-powered election intelligence system for Tamil Nadu using grounded retrieval, production-ready architecture, and reliable political knowledge access.

L

Lugman Hussain Khan

research
Improving long-horizon planning through harness engineering

A technical exploration of improving long-horizon planning reliability through structured harness engineering, evaluation constraints, and resilient orchestration design for advanced AI systems.

L

Lugman Hussain Khan

open-source
Bringing visual and interactive chats for any LLM

How Weave enables browser-native visual and interactive chat experiences for any OpenAI-compatible model through streamed inline HTML rendering.

Ready to make AI part of how your business operates?

Let's identify the workflows where AI can create the greatest value and determine the right way to build them.