---
author: "Umesh Malik"
canonical: "https://umesh-malik.com/author/umesh-malik"
description: "Umesh Malik is a software engineer writing about AI, Claude Code, LLMs, OpenAI, Anthropic, and developer tooling. 121 articles on AI engineering and production systems."
title: "Umesh Malik — Software Engineer, AI Builder & Writer"
tokens: 1166
generator: "scripts/generate-page-markdown.mjs"
---

![Umesh Malik — Software Engineer, AI Builder & Writer](https://umesh-malik.com/images/umesh-malik.jpg)

Author

# Umesh Malik

Software Engineer · AI Builder · Writer · Expedia Group

Software engineer writing about AI, Claude Code, LLMs, OpenAI, Anthropic, and developer tooling. 121 articles on AI engineering, production systems, and the tools shaping modern development. 5+ years at Expedia Group, Tekion, and BYJU'S.

Claude CodeAI Coding AgentsLLM EngineeringRAGFine-tuningOpenAIAnthropicNext.jsSvelteKitDeveloper Tooling

[X(opens in new tab)](https://x.com/lumeshmalik) [LinkedIn(opens in new tab)](https://linkedin.com/in/umesh-malik) [GitHub(opens in new tab)](https://github.com/Umeshmalik) [Email](mailto:ask@umesh-malik.com)

[Full bio →](https://umesh-malik.com/about)

Google Search · Preferred sources

## Prefer this site on Google

If you already read this writing, add umesh-malik.com as a Preferred Source. Google can then highlight it with a preferred badge in Top Stories, AI Overviews, and AI Mode — for you, not as a site-wide ranking boost.

 [Open in Google Search(opens Google Preferred Sources in a new tab)](https://www.google.com/preferences/source?q=umesh-malik.com)

## Recent Articles 121 total

[All articles →](https://umesh-malik.com/blog)

 [![Dashboard-style cover for reducing a Rust struct's memory footprint, showing five layout techniques and a 953-to-420-byte result](https://umesh-malik.com/blog/reduce-rust-struct-memory-footprint-cover.png)

Web Engineering • Sep 4, 2026

### How to Reduce Rust Struct Memory Footprint: 5 Techniques, 56% Smaller

How to reduce Rust struct memory footprint: the 5 layout changes that shrank a real cache entry from 953 to 420 bytes, and what each one costs.

10 min read

Read more →](https://umesh-malik.com/blog/reduce-rust-struct-memory-footprint)

 [![Migration flowchart showing OpenAI Python SDK moving from httpx with certifi to httpx2 with OS trust store](https://umesh-malik.com/blog/openai-python-httpx2-migration-guide-cover.png)

LLM Engineering • Aug 29, 2026

### OpenAI Python HTTPX2 Migration: Fix the TLS Trap First

The OpenAI Python HTTPX2 migration breaks certifi TLS in containers and proxies. The full checklist, OS trust store fix, and legacy escape hatch.

6 min read

Read more →](https://umesh-malik.com/blog/openai-python-httpx2-migration-guide)

 [![Cover showing the PCIe bottleneck in traditional GPU-NIC architecture versus MTIA 300's built-in NIC design that delivers 1.2 TB/s without CPU mediation](https://umesh-malik.com/blog/eliminate-pcie-bottleneck-ai-training-cover.png)

AI Engineering • Aug 25, 2026

### Fix the PCIe Bottleneck in AI Training: How Built-in NICs Work

Fix the PCIe bottleneck in AI training with built-in NICs. Meta's MTIA 300 reclaims 1.2 TB/s by eliminating host CPU mediation.

9 min read

Read more →](https://umesh-malik.com/blog/eliminate-pcie-bottleneck-ai-training)

 [![The flow from prompt to invisible watermark: Paint sends a prompt to Microsoft's moderation server, receives a GUID, runs local inference, then embeds that GUID into the pixels](https://umesh-malik.com/blog/ms-paint-invisible-watermark-guid-cover.png)

AI Security • Aug 25, 2026

### MS Paint Invisible Watermark: How to Find the GUID in AI Images

The MS Paint invisible watermark embeds a server GUID into every AI image — even when inference runs locally. Here's how to detect it.

8 min read

Read more →](https://umesh-malik.com/blog/ms-paint-invisible-watermark-guid)

 [![Cover showing the inference engine attack surface with token stream flowing from model through vulnerable parser to arbitrary code execution, and the defense architecture separating GPU host from token parsing](https://umesh-malik.com/blog/secure-llm-inference-vllm-cve-2025-9141-cover.png)

AI Security • Aug 25, 2026

### How to Harden vLLM Inference: CVE-2025-9141 Defense Guide

How to harden vLLM inference against token exploits. CVE-2025-9141 let models run code via eval(). Separate GPU hosts from parsers.

9 min read

Read more →](https://umesh-malik.com/blog/secure-llm-inference-vllm-cve-2025-9141)

 [![Chart showing ChatGPT Search site-scoped query share jumping from 0.3% to 17% on August 8, 2026](https://umesh-malik.com/blog/chatgpt-search-site-scoping-geo-cover.png)

LLM Engineering • Aug 24, 2026

### ChatGPT Search Optimization After the Site-Scoping Shift

ChatGPT search optimization changed when 17% of queries started scoping to specific sites. What the GPT-5.6 shift means and how to get cited.

7 min read

Read more →](https://umesh-malik.com/blog/chatgpt-search-site-scoping-geo)

[View all 121 articles →](https://umesh-malik.com/blog)
