VectorMesh – Designing a High-Traffic, Secure-by-Design Enterprise AI-RAG Platform

VectorMesh – Designing a High-Traffic, Secure-by-Design Enterprise AI-RAG Platform TL;DR This article explores the architecture and engineering decisions behind VectorMesh, an enterprise-grade AI-RAG (Retrieval-Augmented Generation) platform designed to address the bottlenecks that emerge when AI systems move from experimentation into production. The platform is built around Rust-based microservices, Kubernetes, Istio, Apache APISIX, and end-to-end observability, with a strong focus on scalability, security, performance, and operational visibility. ...

May 12, 2026 · 5 min · Khomkrit

From 3 Seconds to 2 Milliseconds: Escaping the Pandas CPU Bottleneck with NVIDIA cuDF

Tags: Data Engineering, Python, NVIDIA RAPIDS, Performance Optimization The Problem: The Pandas Bottleneck When building data pipelines—whether for AI/RAG preprocessing, algorithmic trading, or standard ETL—Pandas is the undisputed king of tabular data. However, as your dataset scales to millions of rows, Pandas becomes a notorious CPU bottleneck. Waiting seconds (or minutes) for simple aggregations disrupts the development flow and inflates cloud compute costs. Recently, I explored how to bypass this limitation using NVIDIA RAPIDS cuDF, a library that mirrors the Pandas API but executes strictly on the GPU. ...

February 12, 2024 · 3 min · Khomkrit