<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Narendra Kumar Vadapalli | Blog &amp; Newsletter</title><description>Weekly deep-dives into Frontier LLM architectures, NVIDIA Physical AI robotics stacks, vLLM/SGLang serving, and autonomous agent orchestration.</description><link>https://www.narenvadapalli.com/</link><language>en-us</language><item><title>MLA and FP8 MoE Serving: Multi-Head Latent Attention at Scale</title><link>https://www.narenvadapalli.com/blog/multi-head-latent-attention-mla-fp8-moe-serving/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/multi-head-latent-attention-mla-fp8-moe-serving/</guid><description>Demystify Multi-Head Latent Attention (MLA) and low-precision FP8 Mixture-of-Experts serving. Slash KV-cache footprints by 93% on high-throughput GPUs.</description><pubDate>Tue, 08 Sep 2026 00:00:00 GMT</pubDate><category>ai-inference</category><category>mla</category><category>deepseek</category><category>moe</category><category>fp8</category><category>vllm</category><category>sglang</category><category>kv-cache</category><category>gpu</category></item><item><title>Disaggregated Inference: Separating Prefill and Decode Nodes at Scale</title><link>https://www.narenvadapalli.com/blog/disaggregated-inference-separating-prefill-decode-scale/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/disaggregated-inference-separating-prefill-decode-scale/</guid><description>Scale LLM serving with disaggregated inference. Decouple compute-heavy prefill from memory-bound decode nodes to eliminate TTFT/TPOT interference.</description><pubDate>Mon, 07 Sep 2026 00:00:00 GMT</pubDate><category>ai-inference</category><category>disaggregated-inference</category><category>vllm</category><category>sglang</category><category>distributed-systems</category><category>gpu</category><category>llm-serving</category><category>kv-cache</category></item><item><title>The Neural Rendering Matrix: Comparing NeRFs, Instant-NGP, 3D Gaussian Splatting, and NuRec</title><link>https://www.narenvadapalli.com/blog/neural-rendering-matrix-nerfs-instant-ngp-3dgs-nurec-comparison/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/neural-rendering-matrix-nerfs-instant-ngp-3dgs-nurec-comparison/</guid><description>An exhaustive architectural, mathematical, and benchmark comparison across NeRFs, Instant-NGP, 3D Gaussian Splatting, and NVIDIA NuRec.</description><pubDate>Sun, 06 Sep 2026 00:00:00 GMT</pubDate><category>neural-rendering</category><category>nerf</category><category>instant-ngp</category><category>3d-gaussian-splatting</category><category>nvidia</category><category>computer-vision</category><category>physical-ai</category><category>graphics</category></item><item><title>NVIDIA NuRec &amp; Dynamic 3DGS: Photorealistic Digital Twins for Robotics &amp; AV Simulation</title><link>https://www.narenvadapalli.com/blog/nvidia-nurec-dynamic-3dgs-photorealistic-digital-twins/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/nvidia-nurec-dynamic-3dgs-photorealistic-digital-twins/</guid><description>Explore NVIDIA NuRec: turning drive logs into interactive 3D Gaussian digital twins with dynamic actor decomposition and cross-carline virtual sensor rig adaptation.</description><pubDate>Sat, 05 Sep 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>physical-ai</category><category>autonomous-vehicles</category><category>3d-gaussian-splatting</category><category>digital-twins</category><category>simulation</category><category>robotics</category><category>omniverse</category></item><item><title>Claude Fable 5.1 &amp; Claude Mythos 5.1: Anthropic&apos;s Dual Frontier for Enterprise Coding and High-Assurance Research</title><link>https://www.narenvadapalli.com/blog/claude-fable-mythos-5-1-dual-frontier-intelligence/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/claude-fable-mythos-5-1-dual-frontier-intelligence/</guid><description>Explore Anthropic Claude Fable 5.1 and Mythos 5.1: SWE-bench Pro leadership (81.2%), 75% prompt cache discounts, and dual-persona safety architecture.</description><pubDate>Fri, 04 Sep 2026 00:00:00 GMT</pubDate><category>anthropic</category><category>claude-fable</category><category>claude-mythos</category><category>llms</category><category>swe-bench</category><category>agentic-ai</category><category>cybersecurity</category><category>frontier-models</category></item><item><title>OpenAI GPT-6 Astra: Frontier Agentic Intelligence, ARC-AGI-3, and Critical Risk Thresholds</title><link>https://www.narenvadapalli.com/blog/openai-gpt-6-astra-frontier-agentic-intelligence/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/openai-gpt-6-astra-frontier-agentic-intelligence/</guid><description>Explore OpenAI GPT-6 Astra: architecture, ARC-AGI-3 performance, ExploitBench 100% cybersecurity capabilities, and chain-of-thought safety monitoring.</description><pubDate>Thu, 03 Sep 2026 00:00:00 GMT</pubDate><category>openai</category><category>gpt-6-astra</category><category>llms</category><category>reasoning-models</category><category>cybersecurity</category><category>agentic-ai</category><category>arc-agi</category><category>frontier-models</category></item><item><title>The 3D Gaussian Splatting Revolution: Real-Time Differentiable Primitives</title><link>https://www.narenvadapalli.com/blog/3d-gaussian-splatting-revolution-real-time-differentiable-primitives/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/3d-gaussian-splatting-revolution-real-time-differentiable-primitives/</guid><description>Explore the 3D Gaussian Splatting revolution: explicit covariance parameterization, Spherical Harmonics, GPU tile rasterization, and 100+ FPS rendering.</description><pubDate>Wed, 02 Sep 2026 00:00:00 GMT</pubDate><category>3d-gaussian-splatting</category><category>neural-rendering</category><category>cuda</category><category>computer-vision</category><category>physical-ai</category><category>graphics</category><category>deep-learning</category></item><item><title>Accelerating Implicit Fields: Instant-NGP &amp; Multiresolution Hash Grids</title><link>https://www.narenvadapalli.com/blog/accelerating-implicit-fields-instant-ngp-multiresolution-hash-grids/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/accelerating-implicit-fields-instant-ngp-multiresolution-hash-grids/</guid><description>Discover how Instant-NGP eliminated NeRF&apos;s rendering latency by pairing hierarchical multiresolution spatial hash tables with tiny fully-fused CUDA MLPs.</description><pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate><category>instant-ngp</category><category>neural-rendering</category><category>nerf</category><category>cuda</category><category>computer-vision</category><category>physical-ai</category><category>graphics</category><category>deep-learning</category></item><item><title>Demystifying NeRFs: Volumetric Rendering &amp; Implicit Coordinate Networks</title><link>https://www.narenvadapalli.com/blog/demystifying-nerfs-volumetric-rendering-implicit-coordinate-networks/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/demystifying-nerfs-volumetric-rendering-implicit-coordinate-networks/</guid><description>Explore the mechanics of Neural Radiance Fields (NeRFs), from 5D continuous coordinate networks and positional encodings to numerical volumetric quadrature.</description><pubDate>Mon, 31 Aug 2026 00:00:00 GMT</pubDate><category>neural-rendering</category><category>nerf</category><category>implicit-neural-representations</category><category>computer-vision</category><category>physical-ai</category><category>graphics</category><category>deep-learning</category></item><item><title>NVIDIA Drive Cosmos &amp; Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models</title><link>https://www.narenvadapalli.com/blog/nvidia-drive-cosmos-cosmos-drive-dreams-world-foundation-models/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/nvidia-drive-cosmos-cosmos-drive-dreams-world-foundation-models/</guid><description>Explore NVIDIA Drive Cosmos and Cosmos-Drive-Dreams for scalable, multi-view synthetic driving data generation powered by World Foundation Models.</description><pubDate>Sun, 30 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>physical-ai</category><category>autonomous-vehicles</category><category>cosmos</category><category>world-models</category><category>synthetic-data</category><category>diffusion-models</category><category>robotics</category><category>simulation</category></item><item><title>The Evolution of Edge AI: NVIDIA Jetson Nano, Jetson Orin Nano, and the All-New Jetson Orin Nano 2</title><link>https://www.narenvadapalli.com/blog/evolution-of-edge-ai-nvidia-jetson-nano-orin-nano-2/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/evolution-of-edge-ai-nvidia-jetson-nano-orin-nano-2/</guid><description>Explore the generational leap in edge robotics silicon from the original 2019 Jetson Nano to the 78 TOPS Jetson Orin Nano 2 for Physical AI.</description><pubDate>Sat, 29 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>physical-ai</category><category>robotics</category><category>jetson</category><category>edge-ai</category><category>orin-nano-2</category><category>vla</category><category>computer-vision</category><category>embedded-systems</category></item><item><title>NVIDIA NIM: Containerized Enterprise GenAI Serving Architecture</title><link>https://www.narenvadapalli.com/blog/nvidia-nim-containerized-enterprise-genai-serving/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/nvidia-nim-containerized-enterprise-genai-serving/</guid><description>Explore NVIDIA NIM (Inference Microservices), the containerized serving stack integrating TensorRT-LLM, vLLM, and KServe.</description><pubDate>Fri, 28 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>ai-inference</category><category>nim</category><category>kserve</category><category>kubernetes</category><category>tensorrt-llm</category><category>vllm</category><category>enterprise-ai</category><category>docker</category></item><item><title>dynamo-vllm: High-Throughput Distributed PagedAttention at Scale</title><link>https://www.narenvadapalli.com/blog/dynamo-vllm-high-throughput-distributed-pagedattention/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/dynamo-vllm-high-throughput-distributed-pagedattention/</guid><description>Explore dynamo-vllm, combining NVIDIA Dynamo distributed orchestration with vLLM PagedAttention for high-throughput multi-GPU serving.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>ai-inference</category><category>vllm</category><category>dynamo</category><category>pagedattention</category><category>cuda-graphs</category><category>torch-compile</category><category>distributed-systems</category></item><item><title>NVIDIA Triton (Dynamo-Triton): Enterprise Multi-Model Serving Architecture</title><link>https://www.narenvadapalli.com/blog/nvidia-triton-dynamo-triton-enterprise-multi-model-serving/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/nvidia-triton-dynamo-triton-enterprise-multi-model-serving/</guid><description>Explore NVIDIA Triton (Dynamo-Triton), the multi-framework inference server powering concurrent model pipelines and dynamic batching.</description><pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>ai-inference</category><category>triton</category><category>dynamo</category><category>model-serving</category><category>tensorrt</category><category>onnx</category><category>pytorch</category><category>enterprise-ml</category></item><item><title>NVIDIA Dynamo: Data Center-Scale Disaggregated Generative AI Orchestration</title><link>https://www.narenvadapalli.com/blog/nvidia-dynamo-disaggregated-generative-ai-orchestration/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/nvidia-dynamo-disaggregated-generative-ai-orchestration/</guid><description>Explore NVIDIA Dynamo, the distributed inference orchestration platform separating prefill and decode across clusters with smart KV routing.</description><pubDate>Tue, 25 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>ai-inference</category><category>dynamo</category><category>triton</category><category>llm-serving</category><category>disaggregated-inference</category><category>vllm</category><category>sglang</category><category>tensorrt-llm</category></item><item><title>Inside Newton: Open-Source Differentiable Physics for Generalist Robotics</title><link>https://www.narenvadapalli.com/blog/inside-nvidia-newton-differentiable-physics-engine/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/inside-nvidia-newton-differentiable-physics-engine/</guid><description>Explore Newton, the open-source differentiable physics engine co-developed by NVIDIA, Google DeepMind, and Disney Research for robot learning.</description><pubDate>Mon, 24 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>physical-ai</category><category>robotics</category><category>newton</category><category>differentiable-physics</category><category>openusd</category><category>warp</category><category>isaac-lab</category><category>reinforcement-learning</category></item><item><title>Part 9: Demystifying Autonomous Vehicles: The 3-Computer Architecture, SAE Autonomy Levels, and the Sensor Fusion Triad</title><link>https://www.narenvadapalli.com/blog/demystifying-autonomous-vehicles-sae-levels-sensor-fusion/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/demystifying-autonomous-vehicles-sae-levels-sensor-fusion/</guid><description>Explore the 3-computer AV architecture, SAE Levels 0 to 5, the Perception-Planning-Control triad, and multimodal Camera-Radar-LiDAR sensor fusion.</description><pubDate>Sun, 23 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>physical-ai</category><category>autonomous-vehicles</category><category>sensor-fusion</category><category>lidar</category><category>radar</category><category>sae-levels</category><category>robotics</category></item><item><title>Part 8: Silicon at the Edge: NVIDIA Jetson Thor Architecture &amp; Isaac ROS Acceleration</title><link>https://www.narenvadapalli.com/blog/silicon-at-the-edge-nvidia-jetson-thor-isaac-ros/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/silicon-at-the-edge-nvidia-jetson-thor-isaac-ros/</guid><description>Explore NVIDIA Jetson Thor edge computing, Blackwell 800 TFLOPS architecture, NITROS zero-copy IPC, and Isaac ROS sub-50ms humanoid reflexes.</description><pubDate>Sat, 22 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>physical-ai</category><category>robotics</category><category>jetson-thor</category><category>isaac-ros</category><category>ros2</category><category>nitros</category><category>edge-ai</category><category>vslam</category></item><item><title>Part 7: From Simulation to Streets: NVIDIA DRIVE &amp; Alpamayo Autonomous Vehicle Architecture</title><link>https://www.narenvadapalli.com/blog/from-simulation-to-streets-nvidia-drive-alpamayo-av-architecture/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/from-simulation-to-streets-nvidia-drive-alpamayo-av-architecture/</guid><description>Explore NVIDIA DRIVE Thor, Alpamayo AV foundation models, surround-view Bird&apos;s-Eye-View (BEV) transformer fusion, and ASIL-D safety-critical redundancy.</description><pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate><category>nvidia</category><category>physical-ai</category><category>autonomous-vehicles</category><category>drive-thor</category><category>alpamayo</category><category>bevformer</category><category>transformers</category><category>safety-critical</category><category>robotics</category></item><item><title>Part 6: Inside Project GR00T: Vision-Language-Action (VLA) Tokenization &amp; Diffusion Action Heads</title><link>https://www.narenvadapalli.com/blog/inside-project-gr00t-vla-diffusion-heads/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/inside-project-gr00t-vla-diffusion-heads/</guid><description>A deep architectural dissection of NVIDIA Project GR00T: multimodal VLA tokenization, transformer cross-attention backbones, and diffusion policy action heads for humanoid robotics.</description><pubDate>Thu, 20 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>nvidia</category><category>gr00t</category><category>vla</category><category>humanoid-robotics</category><category>physical-ai</category><category>diffusion-policy</category><category>transformers</category><category>architecture</category></item><item><title>Part 5: Scaling Physics with Isaac Sim &amp; Omniverse Replicator: GPU Dynamics, Synthetic Sensors, and Domain Randomization</title><link>https://www.narenvadapalli.com/blog/scaling-physics-isaac-sim-omniverse-replicator/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/scaling-physics-isaac-sim-omniverse-replicator/</guid><description>An in-depth engineering guide to NVIDIA Isaac Sim and Omniverse Replicator: GPU-accelerated PhysX 5 dynamics, synthetic sensor pipelines, and automated domain randomization.</description><pubDate>Wed, 19 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>nvidia</category><category>isaac-sim</category><category>replicator</category><category>robotics</category><category>physical-ai</category><category>simulation</category><category>physx</category><category>architecture</category></item><item><title>Part 4: Demystifying OpenUSD: Architecture, Composition Arcs, usdview, and Simulation Assets</title><link>https://www.narenvadapalli.com/blog/demystifying-openusd-architecture-and-tools/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/demystifying-openusd-architecture-and-tools/</guid><description>A comprehensive guide to OpenUSD (Universal Scene Description): hierarchical scene graphs, LIVRPS composition arcs, step-by-step usdview visualization, and SimReady assets.</description><pubDate>Tue, 18 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>openusd</category><category>nvidia</category><category>omniverse</category><category>usdview</category><category>simready</category><category>digital-twins</category><category>physical-ai</category><category>robotics</category><category>architecture</category></item><item><title>Part 3: Unlocking NVIDIA Omniverse: Architecture, OpenUSD, RTX Rendering, and the Industrial Metaverse Ecosystem</title><link>https://www.narenvadapalli.com/blog/unlocking-nvidia-omniverse-architecture/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/unlocking-nvidia-omniverse-architecture/</guid><description>A comprehensive architectural deep-dive into NVIDIA Omniverse: OpenUSD scene graphs, Nucleus live-sync collaboration, RTX real-time ray tracing, and industrial digital twin ecosystems.</description><pubDate>Mon, 17 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>nvidia</category><category>omniverse</category><category>openusd</category><category>digital-twins</category><category>physical-ai</category><category>robotics</category><category>ray-tracing</category><category>architecture</category></item><item><title>Part 2: Inside NVIDIA Cosmos: World Foundation Models for Physical Commonsense &amp; Video Trajectories</title><link>https://www.narenvadapalli.com/blog/inside-nvidia-cosmos-world-foundation-models/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/inside-nvidia-cosmos-world-foundation-models/</guid><description>A deep architectural dive into NVIDIA Cosmos world foundation models: Mixture-of-Transformers (MoT), continuous spatiotemporal latent tokenizers, and physics-aware trajectory generation.</description><pubDate>Sun, 16 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>nvidia</category><category>cosmos</category><category>world-models</category><category>physical-ai</category><category>robotics</category><category>diffusion</category><category>transformers</category><category>architecture</category></item><item><title>Part 1: Unpacking the NVIDIA Physical AI Data Factory (PAIDF) Stack</title><link>https://www.narenvadapalli.com/blog/unpacking-nvidia-paidf-physical-ai-stack/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/unpacking-nvidia-paidf-physical-ai-stack/</guid><description>An architectural deep-dive into NVIDIA&apos;s Physical AI Data Factory (PAIDF) stack: synthetic data generation with Cosmos &amp; Isaac Sim, VLA foundation models, and edge runtime orchestration.</description><pubDate>Sat, 15 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>nvidia</category><category>physical-ai</category><category>paidf</category><category>robotics</category><category>isaac-sim</category><category>cosmos</category><category>vla</category><category>architecture</category></item><item><title>NVIDIA Nemotron 3.5 Lightning Deep-Dive: 30B MoE Architecture, 3B Active Params, and Local Ollama Execution</title><link>https://www.narenvadapalli.com/blog/nvidia-nemotron-3-5-lightning-architecture-ollama-guide/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/nvidia-nemotron-3-5-lightning-architecture-ollama-guide/</guid><description>How NVIDIA&apos;s Nemotron 3.5 Lightning achieves 3B active parameter speed with 30B MoE capacity, 1M context window, and native Ollama local agent execution.</description><pubDate>Fri, 14 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>nvidia</category><category>nemotron</category><category>moe</category><category>ollama</category><category>local-llm</category><category>agentic-ai</category><category>architecture</category></item><item><title>Frontier MoE Deep-Dive: Analyzing Alibaba&apos;s Qwen 3.8 Flagship Architecture, Performance, and Token Pricing</title><link>https://www.narenvadapalli.com/blog/analyzing-alibabas-qwen-3-8-flagship-moe-model/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/analyzing-alibabas-qwen-3-8-flagship-moe-model/</guid><description>How does Alibaba&apos;s Qwen 3.8 flagship model achieve GPT-5 class reasoning at $0.30 per 1M tokens? Explore 512-expert sparse MoE routing, Multi-Head Latent Attention (MLA), performance benchmarks, and token economics.</description><pubDate>Thu, 13 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>qwen</category><category>alibabacloud</category><category>moe</category><category>deepseek</category><category>llm</category><category>architecture</category><category>pricing</category><category>machine-learning</category></item><item><title>Part 9: The Evolutionary Arc of Computer Vision: From LeNet-5 and ResNet to ConvNeXt and 3D Video Models</title><link>https://www.narenvadapalli.com/blog/evolutionary-arc-computer-vision-lenet-resnet-convnext-3d-video/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/evolutionary-arc-computer-vision-lenet-resnet-convnext-3d-video/</guid><description>How did computer vision evolve from reading bank check zip codes in 1998 to 3D video spatial physics in physical AI and robotics?</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>computer-vision</category><category>cnn</category><category>resnet</category><category>convnext</category><category>3d-video</category><category>spatial-ai</category><category>architecture</category></item><item><title>Part 8: Generative Adversarial Networks (GANs): The Counterfeiter vs. Detective Minimax Game</title><link>https://www.narenvadapalli.com/blog/generative-adversarial-networks-gans-counterfeiter-detective-minimax/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/generative-adversarial-networks-gans-counterfeiter-detective-minimax/</guid><description>Step into the zero-sum game of generative AI—how a counterfeiter Generator and detective Discriminator compete to reach Nash Equilibrium and synthesize photorealistic images.</description><pubDate>Tue, 11 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>gans</category><category>generative-ai</category><category>minimax</category><category>wgan</category><category>pytorch</category><category>architecture</category></item><item><title>Part 7: The Attention Memory Bottleneck: From Self-Attention Basics to MHA, GQA, and DeepSeek&apos;s MLA</title><link>https://www.narenvadapalli.com/blog/attention-memory-bottleneck-mha-gqa-deepseek-mla/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/attention-memory-bottleneck-mha-gqa-deepseek-mla/</guid><description>How does DeepSeek-V3 run 128K context windows with 96.5% less VRAM? Demystify Key-Value Cache growth, MHA, GQA, and low-rank Multi-Head Latent Attention (MLA).</description><pubDate>Mon, 10 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>transformers</category><category>attention</category><category>gqa</category><category>deepseek</category><category>mla</category><category>kv-cache</category><category>architecture</category></item><item><title>Part 6: Why Deep Networks Die: Weight Initialization (He/Xavier), LayerNorm, and Residual Connections</title><link>https://www.narenvadapalli.com/blog/why-deep-networks-die-initialization-layernorm-residual-connections/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/why-deep-networks-die-initialization-layernorm-residual-connections/</guid><description>Why can&apos;t you train a 100-layer neural network without it blowing up or learning nothing, and how do ResNet shortcuts and LayerNorm keep it alive?</description><pubDate>Sun, 09 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>neural-networks</category><category>resnet</category><category>layernorm</category><category>weight-initialization</category><category>optimization</category></item><item><title>Part 5: Inside the Learning Engine: Forward Pass, Backpropagation, and Dynamic Autograd</title><link>https://www.narenvadapalli.com/blog/inside-the-learning-engine-forward-pass-backpropagation-autograd/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/inside-the-learning-engine-forward-pass-backpropagation-autograd/</guid><description>How do neural networks actually learn—how do predictions flow forward, loss errors walk backward via the chain rule, and autograd record every operation?</description><pubDate>Sat, 08 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>neural-networks</category><category>backpropagation</category><category>autograd</category><category>optimization</category><category>pytorch</category></item><item><title>Part 4: Demystifying Activation Functions: Why Neural Networks Need Non-Linearity, Types, and Real-World Use Cases</title><link>https://www.narenvadapalli.com/blog/demystifying-activation-functions-non-linearity-types-use-cases/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/demystifying-activation-functions-non-linearity-types-use-cases/</guid><description>Why are linear models completely useless for complex real-world data, and how do activation functions like ReLU, GELU, and SwiGLU warp space?</description><pubDate>Fri, 07 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>neural-networks</category><category>activation-functions</category><category>relu</category><category>gelu</category><category>architecture</category></item><item><title>Part 3: The Transformer Revolution: How Self-Attention and Q K^T V Solved the GPU Parallelization Bottleneck</title><link>https://www.narenvadapalli.com/blog/transformer-revolution-self-attention-parallelization/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/transformer-revolution-self-attention-parallelization/</guid><description>How did Query-Key-Value self-attention and matrix parallelization eliminate sequential GPU bottlenecks to power ChatGPT?</description><pubDate>Thu, 06 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>neural-networks</category><category>transformer</category><category>self-attention</category><category>architecture</category></item><item><title>Part 2: Why LSTMs Were Needed: Conquering RNN Amnesia, Memory Conveyor Belts, and Gated Doors</title><link>https://www.narenvadapalli.com/blog/why-lstms-were-needed-rnn-amnesia-memory-conveyor-belts-gated-doors/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/why-lstms-were-needed-rnn-amnesia-memory-conveyor-belts-gated-doors/</guid><description>Why do standard RNNs suffer from total amnesia on long sequences, and how did LSTMs solve it with memory conveyor belts and gated doors?</description><pubDate>Wed, 05 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>neural-networks</category><category>lstm</category><category>rnn</category><category>architecture</category></item><item><title>Part 1: Demystifying Neural Networks: From Simple Perceptrons to Deep Neural Networks (DNNs), CNNs, and RNNs</title><link>https://www.narenvadapalli.com/blog/demystifying-neural-networks-perceptron-to-dnn-cnn-rnn/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/demystifying-neural-networks-perceptron-to-dnn-cnn-rnn/</guid><description>How did AI evolve from single-layer perceptrons to deep feedforward networks, 2D spatial CNN filters, and sequential RNN loops?</description><pubDate>Tue, 04 Aug 2026 00:00:00 GMT</pubDate><category>ai</category><category>deep-learning</category><category>neural-networks</category><category>cnn</category><category>rnn</category><category>architecture</category></item><item><title>Google DeepMind&apos;s Gemini Robotics ER 2: The High-Level Brain for Physical AI and Multi-Robot Collaboration</title><link>https://www.narenvadapalli.com/blog/google-deepmind-gemini-robotics-er-2/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/google-deepmind-gemini-robotics-er-2/</guid><description>An architectural deep-dive into Google DeepMind&apos;s Gemini Robotics ER 2 announcement—exploring real-time video streaming, high-level reasoning vs low-level VLA execution, multi-robot team orchestration, and temporal moment-finding.</description><pubDate>Mon, 03 Aug 2026 00:00:00 GMT</pubDate><category>Robotics</category><category>Google DeepMind</category><category>Gemini</category><category>Physical AI</category><category>VLA Models</category><category>Multi-Agent</category><category>AI Architecture</category></item><item><title>Demystifying LoRA (Low-Rank Adaptation): From Training Efficiency to Multi-Adapter Inference</title><link>https://www.narenvadapalli.com/blog/demystifying-lora-low-rank-adaptation/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/demystifying-lora-low-rank-adaptation/</guid><description>A comprehensive developer guide to Low-Rank Adaptation (LoRA). Exploring why LoRA is used, matrix decomposition math, VRAM reductions during training, weight merging vs multi-adapter serving during inference, QLoRA, DoRA, and runnable Python simulations.</description><pubDate>Sun, 02 Aug 2026 00:00:00 GMT</pubDate><category>lora</category><category>qlora</category><category>ai-finetuning</category><category>machine-learning</category><category>vllm</category><category>deep-learning</category><category>architecture</category></item><item><title>Deep-Dive: SGLang v0.5.16 Architecture and High-Throughput Inference Comparison</title><link>https://www.narenvadapalli.com/blog/sglang-v0-5-16-architecture-and-inference-comparison/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/sglang-v0-5-16-architecture-and-inference-comparison/</guid><description>An architectural deep-dive into SGLang v0.5.16. Analyzing RadixAttention KV cache reuse, compressed FSM constrained decoding, Torch Compile CUDA graph optimizations, and multi-engine benchmarks.</description><pubDate>Sat, 01 Aug 2026 00:00:00 GMT</pubDate><category>sglang</category><category>vllm</category><category>ai-inference</category><category>radix-attention</category><category>machine-learning</category><category>architecture</category></item><item><title>Understanding Mixture-of-Experts (MoE): From Specialist Clinics to Kimi K3&apos;s 896-Expert Router</title><link>https://www.narenvadapalli.com/blog/understanding-mixture-of-experts-moe/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/understanding-mixture-of-experts-moe/</guid><description>A developer-friendly guide to Mixture-of-Experts (MoE) architectures. Exploring the specialist clinic analogy, total vs active parameters, Top-K gating networks, expert collapse traps, and how Kimi K3 scales to 896 micro-experts.</description><pubDate>Fri, 31 Jul 2026 00:00:00 GMT</pubDate><category>moe</category><category>ai-architecture</category><category>kimi-k3</category><category>machine-learning</category><category>deep-learning</category><category>llm</category></item><item><title>Hosting Moonshot AI&apos;s Kimi K3 Open Weights with vLLM: High-Throughput Serving at Scale</title><link>https://www.narenvadapalli.com/blog/hosting-kimi-k3-vllm/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/hosting-kimi-k3-vllm/</guid><description>A comprehensive developer guide to hosting Moonshot AI&apos;s Kimi K3 open weights on vLLM. Exploring MXFP4 MoE serving, KDA hybrid prefix caching, DSpark speculative decoding (370 tok/s), and NVIDIA/AMD multi-GPU cluster recipes.</description><pubDate>Thu, 30 Jul 2026 00:00:00 GMT</pubDate><category>kimi-k3</category><category>vllm</category><category>ai-inference</category><category>open-weights</category><category>machine-learning</category><category>architecture</category></item><item><title>Physical AI Models: Grounding Intelligence in Space, Physics, and Robotics</title><link>https://www.narenvadapalli.com/blog/physical-ai-models-grounding-in-space-and-robotics/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/physical-ai-models-grounding-in-space-and-robotics/</guid><description>A technical deep-dive into Physical AI models. Exploring spatial understanding, 3D geometry priors, physics simulation engines, embodied VLA control loops, and robotics integration.</description><pubDate>Wed, 29 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>physical-ai</category><category>robotics</category><category>world-models</category><category>machine-learning</category><category>architecture</category></item><item><title>The DeepSeek Architectural Inflection Point: From MLA to Emergence of Open-Weight Reasoning</title><link>https://www.narenvadapalli.com/blog/deepseek-architectural-inflection-point/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/deepseek-architectural-inflection-point/</guid><description>An architectural deep-dive into DeepSeek&apos;s evolution from DeepSeek-LLM to R1. Exploring Multi-Head Latent Attention (MLA), DeepSeekMoE, MTP, pure RL emergence, and comparisons with contemporary frontier models.</description><pubDate>Tue, 28 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>deepseek</category><category>models</category><category>open-source</category><category>machine-learning</category><category>architecture</category></item><item><title>Scale and Performance: Serving LLMs with vLLM and llm-d</title><link>https://www.narenvadapalli.com/blog/serving-llms-with-vllm-and-llm-d/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/serving-llms-with-vllm-and-llm-d/</guid><description>A comprehensive developer guide to scaling LLM serving using vLLM and llm-d. Explore PagedAttention, continuous batching, disaggregated prefill/decode, and Kubernetes deployment scripts.</description><pubDate>Mon, 27 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>inference</category><category>vllm</category><category>llm-d</category><category>kubernetes</category><category>deep-learning</category></item><item><title>Under the Hood of Moonshot AI&apos;s Kimi K3: The Architecture of 3-Trillion Parameter Thinking Models</title><link>https://www.narenvadapalli.com/blog/moonshot-ai-kimi-k3-thinking-models/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/moonshot-ai-kimi-k3-thinking-models/</guid><description>A comprehensive developer guide to Moonshot AI&apos;s Kimi K3. Comparing K3, K2.7 Code, K2.6, and K2.5 across Preserved Thinking, reasoning effort, 1M context, API quickstart, organizational best practices, and prompt engineering.</description><pubDate>Sun, 26 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>models</category><category>open-source</category><category>thinking-models</category><category>kimi-k3</category><category>prompt-engineering</category></item><item><title>Anthropic&apos;s Claude Opus 5: Frontier Reasoning, Benchmarks, and Prompt Engineering</title><link>https://www.narenvadapalli.com/blog/anthropic-claude-opus-5-architectural-guide/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/anthropic-claude-opus-5-architectural-guide/</guid><description>A deep-dive into Anthropic&apos;s flagship Claude Opus 5 release. Comparing Opus 3 vs. Opus 3.5 vs. Opus 5, analyzing benchmarks, and mastering official prompt engineering techniques.</description><pubDate>Sat, 25 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>claude</category><category>anthropic</category><category>llm</category><category>benchmarks</category><category>prompt-engineering</category></item><item><title>The Architectural Spectrum of World Foundation Models: Renderers, State Simulators, and Action Planners</title><link>https://www.narenvadapalli.com/blog/architecture-of-world-foundation-models/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/architecture-of-world-foundation-models/</guid><description>An engineering breakdown of World Foundation Models—moving beyond text tokens into spatial intelligence, persistent 3D state representations, and perception-action control loops.</description><pubDate>Fri, 24 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>world-models</category><category>robotics</category><category>spatial-ai</category><category>architecture</category><category>software-engineering</category></item><item><title>Google&apos;s Gemini 3 Family: The Comprehensive Developer Guide and Model Comparison</title><link>https://www.narenvadapalli.com/blog/gemini-3-model-family-comparison-guide/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/gemini-3-model-family-comparison-guide/</guid><description>An in-depth guide to Google&apos;s newly expanded Gemini 3 model family—including Gemini 3.6 Flash, 3.5 Flash-Lite, 3.5 Flash Cyber, and 3.1 Pro. Learn how they compare on latency, cost, and capabilities.</description><pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>gemini</category><category>llm</category><category>google</category><category>cloud</category><category>engineering</category></item><item><title>Connecting Telegram to OpenClaw: A Complete Step-by-Step Guide</title><link>https://www.narenvadapalli.com/blog/openclaw-telegram-step-by-step-guide/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/openclaw-telegram-step-by-step-guide/</guid><description>Learn how to connect OpenClaw to Telegram from scratch. A beginner-friendly step-by-step tutorial using BotFather to run your AI butler locally.</description><pubDate>Wed, 22 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agents</category><category>open-source</category><category>automation</category><category>telegram</category></item><item><title>LangChain vs. LangGraph: Moving from Chains to Cyclic State Graphs</title><link>https://www.narenvadapalli.com/blog/langchain-vs-langgraph-cyclic-state-graphs/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/langchain-vs-langgraph-cyclic-state-graphs/</guid><description>Why linear chains fail at complex agentic workflows. Learn how LangGraph models agents as state machines using persistent schemas, nodes, and cyclic edges.</description><pubDate>Tue, 21 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agents</category><category>langchain</category><category>langgraph</category><category>software-engineering</category></item><item><title>vLLM vs. llama.cpp: Which is the Real Production King?</title><link>https://www.narenvadapalli.com/blog/vllm-vs-llamacpp-production-comparison/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/vllm-vs-llamacpp-production-comparison/</guid><description>A deep-dive architectural comparison between vLLM and llama.cpp. Compare throughput, memory footprints, and choose the right engine for your workload.</description><pubDate>Mon, 20 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>infrastructure</category><category>performance</category><category>software-engineering</category></item><item><title>Thinking Machines&apos; Inkling: Under the Hood of the 975B Parameter Open Multimodal MoE</title><link>https://www.narenvadapalli.com/blog/thinking-machines-inkling-open-multimodal-moe/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/thinking-machines-inkling-open-multimodal-moe/</guid><description>A deep dive into Mira Murati&apos;s new open-weights multimodal model, Inkling. Analyze the 975B MoE architecture, active parameter routing, and local serving.</description><pubDate>Sun, 19 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>models</category><category>open-source</category><category>architecture</category></item><item><title>Token Economics, LLM Gateways, and Router9</title><link>https://www.narenvadapalli.com/blog/token-economics-llm-gateways-router9/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/token-economics-llm-gateways-router9/</guid><description>Understand the financial side of LLM serving. Calculate cost-per-token, setup open-source LLM gateways, and implement intelligent routers like Router9.</description><pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>architecture</category><category>finops</category><category>software-engineering</category></item><item><title>Inference Optimizations: Speeding up Prefill and Decode</title><link>https://www.narenvadapalli.com/blog/inference-optimizations-prefill-decode/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/inference-optimizations-prefill-decode/</guid><description>Explore advanced LLM serving optimizations. Learn how FlashAttention, chunked prefill, speculative decoding, and KV cache eviction speed up TTFT and inter-token latency.</description><pubDate>Fri, 17 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>machine-learning</category><category>performance</category><category>infrastructure</category></item><item><title>OpenClaw in Action: Connecting WhatsApp to Automated Workflows</title><link>https://www.narenvadapalli.com/blog/openclaw-whatsapp-workflows/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/openclaw-whatsapp-workflows/</guid><description>A step-by-step developer guide to configuring OpenClaw&apos;s WhatsApp gateway, setting up secure session authorization, and building automated task sync skills.</description><pubDate>Thu, 16 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agents</category><category>openclaw</category><category>whatsapp</category><category>notion</category><category>automation</category></item><item><title>The Self-Hosted AI Butler: Modular Assistance with OpenClaw</title><link>https://www.narenvadapalli.com/blog/openclaw-self-hosted-ai-butler/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/openclaw-self-hosted-ai-butler/</guid><description>Learn how to build and host your own AI assistant with OpenClaw. Configure the ClawHub modular skill store, write SKILL.md files, and connect to chat platforms.</description><pubDate>Wed, 15 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agents</category><category>openclaw</category><category>productivity</category><category>self-hosted</category></item><item><title>Nous Research&apos;s Hermes Agent: Under the Hood</title><link>https://www.narenvadapalli.com/blog/hermes-agent-self-improving-systems/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/hermes-agent-self-improving-systems/</guid><description>Deep-dive into the configuration, CLI commands, and skill architecture of Nous Research&apos;s Hermes Agent. Learn to manage tool gateways and design custom skills.</description><pubDate>Tue, 14 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agents</category><category>hermes-agent</category><category>software-engineering</category></item><item><title>The Landscape of Agentic AI: From Single-Agent Scripts to Multi-Agent Networks</title><link>https://www.narenvadapalli.com/blog/landscape-of-agentic-ai/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/landscape-of-agentic-ai/</guid><description>Start the Autonomous AI Agents Series. Learn about the ReAct pattern, single-agent context decay, and multi-agent coordination graph topologies.</description><pubDate>Mon, 13 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agents</category><category>software-engineering</category><category>architecture</category></item><item><title>The Landscape of LLM Inference Engines: Open Source vs. Enterprise</title><link>https://www.narenvadapalli.com/blog/inference-engines-landscape/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/inference-engines-landscape/</guid><description>A comprehensive developer guide comparing vLLM, TensorRT-LLM, TGI, SGLang, and llama.cpp. Learn about PagedAttention, RadixAttention, and GGUF.</description><pubDate>Sun, 12 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>machine-learning</category><category>cloud-computing</category><category>infrastructure</category></item><item><title>Understanding the KV Cache: The VRAM Bottleneck of LLM Serving</title><link>https://www.narenvadapalli.com/blog/understanding-kv-cache/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/understanding-kv-cache/</guid><description>Why do LLM serving nodes run out of VRAM? Explore the mechanics, calculations, and capacity bottlenecks of the Key-Value (KV) Cache.</description><pubDate>Sat, 11 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>machine-learning</category><category>performance</category><category>infrastructure</category></item><item><title>The Two Pillars of LLM Inference: Prefill vs. Decode</title><link>https://www.narenvadapalli.com/blog/prefill-vs-decode/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/prefill-vs-decode/</guid><description>Explore the two execution phases of LLM inference: compute-bound prefill vs. memory-bandwidth-bound decode. Understand arithmetic intensity and the role of the KV cache.</description><pubDate>Fri, 10 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>machine-learning</category><category>systems-engineering</category><category>performance</category></item><item><title>Basics of AI Inference: Demystifying Latency, Throughput, and Serving</title><link>https://www.narenvadapalli.com/blog/basics-of-ai-inference/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/basics-of-ai-inference/</guid><description>Start the AI Inference Deep-Dive Series. Learn the fundamentals of inference vs training, key performance metrics, and the systems engineering behind serving.</description><pubDate>Thu, 09 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>machine-learning</category><category>engineering</category><category>architecture</category></item><item><title>The Model Taxonomy: LLMs, Vision Models, VLAs, and Diffusion</title><link>https://www.narenvadapalli.com/blog/model-taxonomy/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/model-taxonomy/</guid><description>A developer&apos;s guide to the modern AI model taxonomy. Understand the architecture, modalities, and performance profiles of LLMs, ViTs, VLAs, and Diffusion.</description><pubDate>Wed, 08 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>machine-learning</category><category>architecture</category><category>deep-learning</category></item><item><title>Training vs. Inference Lifecycle: A Developer&apos;s Guide to Weights, Backpropagation, and Serving</title><link>https://www.narenvadapalli.com/blog/training-vs-inference-lifecycle/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/training-vs-inference-lifecycle/</guid><description>A step-by-step developer&apos;s view of the AI model lifecycle—from raw weights and backpropagation during training to freezing weights and serving inference.</description><pubDate>Tue, 07 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>machine-learning</category><category>engineering</category><category>architecture</category></item><item><title>What is a Model Weight? Demystifying Tensors, Matrices, and File Formats</title><link>https://www.narenvadapalli.com/blog/what-is-a-model-weight/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/what-is-a-model-weight/</guid><description>What actually happens when you download a model from Hugging Face? A developer&apos;s guide to tensors, weights, biases, and serialization formats like Safetensors and GGUF.</description><pubDate>Mon, 06 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>machine-learning</category><category>engineering</category><category>basics</category></item><item><title>Claude Code Custom Plugins: Build and Host Your Own Marketplace</title><link>https://www.narenvadapalli.com/blog/claude-code-custom-plugin-marketplace/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/claude-code-custom-plugin-marketplace/</guid><description>Learn how to build, structure, and host a custom, decentralized Claude Code plugin marketplace using a simple GitHub repository.</description><pubDate>Sun, 05 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>claude</category><category>claude-code</category><category>customization</category></item><item><title>Running Local LLMs: Ollama vs. vLLM</title><link>https://www.narenvadapalli.com/blog/running-local-llms-ollama-vllm/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/running-local-llms-ollama-vllm/</guid><description>A comprehensive guide to running Large Language Models locally on Windows, macOS, and Linux using Ollama and vLLM, including pros, cons, and performance tuning.</description><pubDate>Sat, 04 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>llm</category><category>local-inference</category><category>ollama</category><category>vllm</category></item><item><title>Anthropic&apos;s Claude Model Family: Specs, Pros, Cons, and Use Cases</title><link>https://www.narenvadapalli.com/blog/claude-models-comparison-guide/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/claude-models-comparison-guide/</guid><description>Compare Anthropic&apos;s Claude models—Fable 5, Opus 4.8, Sonnet 5, Haiku 4.5, and Claude Science. Learn their pros, cons, and the best use cases.</description><pubDate>Fri, 03 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>claude</category><category>llm</category></item><item><title>Anthropic&apos;s Mid-2026 Wave: Claude Sonnet 5, Claude Science, and Fable 5 Redeployment</title><link>https://www.narenvadapalli.com/blog/claude-sonnet-5-science-workbench-fable-redeployed/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/claude-sonnet-5-science-workbench-fable-redeployed/</guid><description>A summary of Anthropic&apos;s major announcements: the launch of Claude Sonnet 5, the new Claude Science AI workbench, and the global redeployment of Fable 5.</description><pubDate>Thu, 02 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>claude</category><category>announcement</category></item><item><title>Claude Code Custom Skills: Design Methodology and Workspace Personas</title><link>https://www.narenvadapalli.com/blog/claude-code-custom-skills/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/claude-code-custom-skills/</guid><description>Learn how to design, structure, and implement persistent, context-aware custom agent skills in Claude Code to automate code reviews and git workflows.</description><pubDate>Wed, 01 Jul 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>claude</category><category>claude-code</category><category>customization</category></item><item><title>Claude Code Basics: Commands, Subagents, and Memory Layers</title><link>https://www.narenvadapalli.com/blog/claude-code-commands-agents-memory/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/claude-code-commands-agents-memory/</guid><description>Learn the difference between slash commands and file context mentions, configure global vs local memory, and understand how custom skills and subagents operate in Claude Code.</description><pubDate>Tue, 30 Jun 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>claude</category><category>claude-code</category><category>customization</category></item><item><title>Claude Code Customization: CLAUDE.md, AGENTS.md, and SKILLS.md</title><link>https://www.narenvadapalli.com/blog/claude-code-special-files/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/claude-code-special-files/</guid><description>Master project-level customization for Anthropic&apos;s Claude Code by writing rules and guidelines in CLAUDE.md, AGENTS.md, and custom SKILLS.md.</description><pubDate>Mon, 29 Jun 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>claude</category><category>claude-code</category><category>customization</category></item><item><title>Setting up Claude Code: The Ultimate Terminal AI Pair Programmer</title><link>https://www.narenvadapalli.com/blog/setting-up-claude-code/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/setting-up-claude-code/</guid><description>A step-by-step setup guide for installing, configuring, and authenticating Anthropic&apos;s Claude Code CLI across macOS, Windows, and Linux.</description><pubDate>Sun, 28 Jun 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>claude</category><category>claude-code</category></item><item><title>LLMs, Agents, and Harnesses: Demystifying Claude Code</title><link>https://www.narenvadapalli.com/blog/intro-to-claude-code/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/intro-to-claude-code/</guid><description>Confused about the difference between Claude Code and Claude Opus? A deep dive into the architecture of modern agentic coding tools—from Large Language Models (LLMs) to Autonomous Agents and Execution Harnesses.</description><pubDate>Sat, 27 Jun 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>claude</category><category>claude-code</category></item><item><title>Custom Model Endpoints: Hooking up Local &amp; Enterprise Models with Antigravity CLI</title><link>https://www.narenvadapalli.com/blog/antigravity-cli-custom-endpoints/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/antigravity-cli-custom-endpoints/</guid><description>Learn how to configure the Google Antigravity CLI (agy) to run with local LLMs (Ollama, vLLM) or private enterprise endpoints for maximum privacy and cost efficiency.</description><pubDate>Wed, 24 Jun 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>antigravity</category></item><item><title>Going Async: Background Tasks, Timers, and Scheduling in Antigravity CLI</title><link>https://www.narenvadapalli.com/blog/antigravity-cli-background-tasks/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/antigravity-cli-background-tasks/</guid><description>A deep dive into managing long-running tasks, using asynchronous tools, setting timers, and configuring recurring jobs with the Google Antigravity CLI.</description><pubDate>Tue, 23 Jun 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>antigravity</category></item><item><title>Extending Antigravity CLI: Building Custom Skills and Project Rules</title><link>https://www.narenvadapalli.com/blog/extending-antigravity-cli-with-custom-skills/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/extending-antigravity-cli-with-custom-skills/</guid><description>Learn how to supercharge the Google Antigravity CLI (agy) by building custom agent skills, project-scoped rules, and custom scripts to automate repository-specific workflows.</description><pubDate>Sun, 21 Jun 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>antigravity</category></item><item><title>Setting up Google Antigravity CLI: The AI-First Terminal Assistant</title><link>https://www.narenvadapalli.com/blog/setting-up-antigravity-cli/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/setting-up-antigravity-cli/</guid><description>A comprehensive guide on installing and configuring Google Antigravity CLI (agy) on Mac, Windows, and Linux, complete with TUI commands, settings.json customization, and real-world usage examples.</description><pubDate>Sat, 20 Jun 2026 00:00:00 GMT</pubDate><category>ai</category><category>agentic</category><category>antigravity</category></item><item><title>Python 3.15: A First Look at the Coolest New Features</title><link>https://www.narenvadapalli.com/blog/whats-new-in-python-3-15/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/whats-new-in-python-3-15/</guid><description>A deep dive into the most exciting developer features in the Python 3.15 pre-release, including explicit lazy imports, frozendict, built-in sentinels, unpacking in comprehensions, and the Tachyon profiler.</description><pubDate>Fri, 19 Jun 2026 00:00:00 GMT</pubDate><category>python</category><category>programming</category><category>tech</category></item><item><title>SLURM Demo on AWS Ubuntu EC2 instance</title><link>https://www.narenvadapalli.com/blog/slurm-on-aws-ubuntu-ec2/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/slurm-on-aws-ubuntu-ec2/</guid><description>Demo of slurm usage on a single instance of Ubuntu 24.04 EC2 instances on AWS</description><pubDate>Wed, 29 May 2024 00:00:00 GMT</pubDate><category>hpc</category><category>slurm</category><category>aws</category><category>ubuntu</category><category>tech</category></item><item><title>SLURM on WSL</title><link>https://www.narenvadapalli.com/blog/slurm-on-wsl/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/slurm-on-wsl/</guid><description>Setting up SLURM on WSL</description><pubDate>Mon, 27 May 2024 00:00:00 GMT</pubDate><category>hpc</category><category>slurm</category><category>wsl</category><category>windows</category><category>tech</category></item><item><title>Introduction to SLURM</title><link>https://www.narenvadapalli.com/blog/slurm-intro/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/slurm-intro/</guid><description>Simple Linux Utility for Resource Management (SLURM)</description><pubDate>Sun, 26 May 2024 00:00:00 GMT</pubDate><category>hpc</category><category>slurm</category><category>linux</category><category>tech</category></item><item><title>How to find a linux machine is a VM (Virtual Machine) or a Bare Metal</title><link>https://www.narenvadapalli.com/blog/finding-linux-machine-vm-or-baremetal/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/finding-linux-machine-vm-or-baremetal/</guid><description>If you can SSH into a linux machine and want to find out if its baremetal or Virtual Machine</description><pubDate>Tue, 07 Nov 2023 00:00:00 GMT</pubDate><category>linux</category><category>vm</category><category>sysadmin</category><category>tech</category></item><item><title>Storing Github access token in git credential store</title><link>https://www.narenvadapalli.com/blog/github-credentials-store-with-access-token/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/github-credentials-store-with-access-token/</guid><description>Using git credentials store the github access token to avoid the re-prompting of username and pwd</description><pubDate>Tue, 04 Apr 2023 00:00:00 GMT</pubDate><category>github</category><category>git</category><category>security</category><category>tech</category></item><item><title>Token generation for Registering Self Hosted Github Runner via REST API</title><link>https://www.narenvadapalli.com/blog/generating-token-for-github-runner-registration/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/generating-token-for-github-runner-registration/</guid><description>Explains how to generate a token using github API to be used in turn with Github self hosted runner registration</description><pubDate>Tue, 21 Mar 2023 00:00:00 GMT</pubDate><category>github</category><category>ci/cd</category><category>devops</category><category>tech</category></item><item><title>Setting up a Self Hosted Github Runner</title><link>https://www.narenvadapalli.com/blog/self-hosted-github-runner-registration-process/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/self-hosted-github-runner-registration-process/</guid><description>Explains how to setup a Github self hosted runner and register</description><pubDate>Mon, 20 Mar 2023 00:00:00 GMT</pubDate><category>github</category><category>ci/cd</category><category>devops</category><category>tech</category></item><item><title>Managing the NodeJS versions on Windows</title><link>https://www.narenvadapalli.com/blog/node-version-manager-nvm/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/node-version-manager-nvm/</guid><description>Node Version Manager (nvm) helps in managing multiple NodeJS versions</description><pubDate>Sun, 13 Nov 2022 00:00:00 GMT</pubDate><category>nodejs</category><category>nvm</category><category>javascript</category><category>tech</category></item><item><title>Customizing the Powershell terminal with oh-my-posh</title><link>https://www.narenvadapalli.com/blog/customizing-powershell-terminal-oh-my-posh/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/customizing-powershell-terminal-oh-my-posh/</guid><description>Instructions on customizing the terminal in powershell with oh-my-posh and winget</description><pubDate>Thu, 07 Jul 2022 00:00:00 GMT</pubDate><category>powershell</category><category>terminal</category><category>tech</category></item><item><title>File permissions on Windows - chmod 400 in Powershell</title><link>https://www.narenvadapalli.com/blog/chmod-400-windows-powershel/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/chmod-400-windows-powershel/</guid><description>Powershell equivalent of chmod 400 for Windows files</description><pubDate>Thu, 09 Jun 2022 00:00:00 GMT</pubDate><category>windows</category><category>powershell</category><category>security</category><category>tech</category></item><item><title>Github login using access token via command line</title><link>https://www.narenvadapalli.com/blog/github-login-using-access-token-via-cmdline/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/github-login-using-access-token-via-cmdline/</guid><description>Logging in using github access token (no more passwords)</description><pubDate>Wed, 29 Sep 2021 00:00:00 GMT</pubDate><category>github</category><category>git</category><category>cli</category><category>tech</category></item><item><title>Adding Google Analytics to NuxtJS app</title><link>https://www.narenvadapalli.com/blog/google-analytics-ga4-property-nuxtjs-app/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/google-analytics-ga4-property-nuxtjs-app/</guid><description>Adding Google Analytics GA4 property to NuxtJS App</description><pubDate>Thu, 02 Sep 2021 00:00:00 GMT</pubDate><category>nuxtjs</category><category>analytics</category><category>seo</category><category>tech</category></item><item><title>Productive Taskbar Settings missing in Windows 11</title><link>https://www.narenvadapalli.com/blog/important-taskbar-settings-gone-in-win11/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/important-taskbar-settings-gone-in-win11/</guid><description>Very useful Taskbar Settings goes missing in Windows 11.</description><pubDate>Tue, 06 Jul 2021 00:00:00 GMT</pubDate><category>windows</category><category>win11</category><category>tech</category></item><item><title>Secureboot + Ubuntu + VirtualBox Signing kernel modules</title><link>https://www.narenvadapalli.com/blog/virtualbox-ubuntu-secureboot-issue/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/virtualbox-ubuntu-secureboot-issue/</guid><description>Set of steps required for dealing with secureboot on Ubuntu where VirutalBox service has issues</description><pubDate>Sun, 09 May 2021 00:00:00 GMT</pubDate><category>virtualbox</category><category>ubuntu</category><category>secureboot</category><category>tech</category></item><item><title>Fixing the postfix error dpkg</title><link>https://www.narenvadapalli.com/blog/postfix-error-ubuntu-dpkg/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/postfix-error-ubuntu-dpkg/</guid><description>Steps to fix the postfix error happening during apt upgrade ubuntu.</description><pubDate>Wed, 21 Apr 2021 00:00:00 GMT</pubDate><category>linux</category><category>ubuntu</category><category>postfix</category><category>tech</category></item><item><title>Running a react app on Local Kubernetes cluster on Windows 10</title><link>https://www.narenvadapalli.com/blog/running-react-app-on-local-k8s-on-windows/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/running-react-app-on-local-k8s-on-windows/</guid><description>Process and steps for running react app on local k8s cluster using minikube on windows 10</description><pubDate>Tue, 16 Mar 2021 00:00:00 GMT</pubDate><category>react</category><category>k8s</category><category>windows</category><category>docker</category><category>tech</category></item><item><title>Gatsby site hosted on AWS Amplify redirecting to homepage always</title><link>https://www.narenvadapalli.com/blog/gatsby-site-redirecting-to-homepage-aws-amplify/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/gatsby-site-redirecting-to-homepage-aws-amplify/</guid><description>Using the rewrites and redirects on AWS Amplify for the depolyed personal website</description><pubDate>Mon, 02 Nov 2020 00:00:00 GMT</pubDate><category>gatsby</category><category>aws</category><category>amplify</category><category>tech</category></item><item><title>Connecting AWS Amplify for deployment of website</title><link>https://www.narenvadapalli.com/blog/connecting-aws-amplify-for-deployment/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/connecting-aws-amplify-for-deployment/</guid><description>Explains how to connect the gatsby website hosted on github to AWS Amplify for deployment</description><pubDate>Sun, 01 Nov 2020 00:00:00 GMT</pubDate><category>aws</category><category>amplify</category><category>deployment</category><category>tech</category></item><item><title>Evolution of this website</title><link>https://www.narenvadapalli.com/blog/evolution-of-this-website/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/evolution-of-this-website/</guid><description>will be capturing the evolution of my website chronologically (latest first)</description><pubDate>Sat, 31 Oct 2020 00:00:00 GMT</pubDate><category>webdev</category><category>website</category><category>tech</category></item><item><title>Making of &quot;Tree Story&quot; Animated Short</title><link>https://www.narenvadapalli.com/blog/making-of-tree-story/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/making-of-tree-story/</guid><description>Behind the scenes of the Animated Short &quot;Tree Story&quot;</description><pubDate>Wed, 21 Oct 2020 00:00:00 GMT</pubDate><category>cg</category><category>art</category><category>tech</category></item><item><title>Profiling &amp; Visualization Tools in Python - Part 1</title><link>https://www.narenvadapalli.com/blog/profiling-visualization-tools-in-python-part-1/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/profiling-visualization-tools-in-python-part-1/</guid><description>Profiling &amp; Visualization Tools in Python using cProfile</description><pubDate>Mon, 19 Oct 2020 00:00:00 GMT</pubDate><category>python</category><category>profiling</category><category>tech</category></item><item><title>Missing permissions issue linking google analytics and adsense</title><link>https://www.narenvadapalli.com/blog/missing-permissions-linking-google-analytics-adsense/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/missing-permissions-linking-google-analytics-adsense/</guid><description>Problem of linking google analytics and adsense</description><pubDate>Fri, 16 Oct 2020 00:00:00 GMT</pubDate><category>analytics</category><category>adsense</category><category>tech</category></item><item><title>Adding Google Analytics to personal website</title><link>https://www.narenvadapalli.com/blog/google-analytics-to-gatsby-app/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/google-analytics-to-gatsby-app/</guid><description>How I added google analytics to my personal website written in Gatsby</description><pubDate>Wed, 07 Oct 2020 00:00:00 GMT</pubDate><category>gatsby</category><category>analytics</category><category>seo</category><category>tech</category></item><item><title>Chrome Remote Desktop on Fedora</title><link>https://www.narenvadapalli.com/blog/chrome-remote-desktop-on-fedora/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/chrome-remote-desktop-on-fedora/</guid><description>How to install Chrome Remote Desktop on Fedora</description><pubDate>Fri, 27 Mar 2020 00:00:00 GMT</pubDate><category>linux</category><category>fedora</category><category>remote-desktop</category><category>tech</category></item><item><title>Building grpc whl with Maya 2019 on Windows</title><link>https://www.narenvadapalli.com/blog/building-grpc-whl-with-maya-2019-on-win64/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/building-grpc-whl-with-maya-2019-on-win64/</guid><description>My stuggle with building a grpc python library with Maya 2019 on Windows</description><pubDate>Fri, 06 Mar 2020 00:00:00 GMT</pubDate><category>maya</category><category>grpc</category><category>c++</category><category>tech</category></item><item><title>Tech Stack of the website</title><link>https://www.narenvadapalli.com/blog/tech-stack-of-the-website/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/tech-stack-of-the-website/</guid><description>Brief about the tech stack used for this website</description><pubDate>Wed, 08 Jan 2020 00:00:00 GMT</pubDate><category>webdev</category><category>website</category><category>tech</category></item><item><title>Deleting linux from dual boot</title><link>https://www.narenvadapalli.com/blog/deleting-linux-from-dual-boot/</link><guid isPermaLink="true">https://www.narenvadapalli.com/blog/deleting-linux-from-dual-boot/</guid><description>Steps for safely removing linux from dual boot.</description><pubDate>Thu, 02 Jan 2020 00:00:00 GMT</pubDate><category>linux</category><category>windows</category><category>dual-boot</category><category>tech</category></item></channel></rss>