Skip to content
The AppScale ArchiveWriting on Sovereign AI
21.4934° N / 86.9135° EEST. 2025 — India
317+ Essays · 27 Series
Scroll ↓

OldKnowledge,New Vessel.

Essays on on-device AI, data sovereignty, and building systems that keep knowledge where it belongs.

Enter the archive
— 00 / ThesisEvery business runs on knowledge older than its software.
01The Archive

Latest Entries

Full index — 317 essays →
02The Library
श्रीगणेशाय नमः ॥
अथ प्रथमोऽध्यायः ॥
विद्या ददाति विनयं विनयाद् याति पात्रताम् ।
पात्रत्वाद्धनमाप्नोति धनाद्धर्मं ततः सुखम् ॥
[ Coming Soon ]

We digitize centuries-old manuscripts. Then we build with the same discipline.

AppScale's roots are in a quiet project: structuring classical Sanskrit texts into faithful digital form. Extraction, structure, provenance, sovereignty — the same principles now power our client work.

न हि ज्ञानेन सदृशं पवित्रमिह विद्यते ।
“Nothing in this world purifies like knowledge.”
Bhagavad Gita · 4.38
03Capabilities

Built for Every Business

Executive AI Series · MCP Security · 日本語 · Edge AI Engineering ·
RAG in Production · Sovereign AI · Fine-Tuning · Agentic Systems ·
19+ yrs engineering·npm — react-native-edge-vector-store·Read in IN · JP · SG·AppScale LLP — DPIIT recognized
05Full Index
Securing Edge AI in 2026: Defending On-Device Models Against Theft, Tampering, and Extraction
ai-architecture1 min read

Securing Edge AI in 2026: Defending On-Device Models Against Theft, Tampering, and Extraction

You moved inference on-device for speed and privacy — and shipped your million-dollar weights to the attacker. Securing edge AI: theft, tampering, extraction, and the hybrid line.

July 20, 2026Read
Agentic AI in Cybersecurity: Strengthening Defenses Autonomously — Without Handing the Attacker Your Keys
ai-architecture1 min read

Agentic AI in Cybersecurity: Strengthening Defenses Autonomously — Without Handing the Attacker Your Keys

An AI agent can contain an intrusion in 90 seconds — or be turned against you by one poisoned alert. The agentic SOC: bounded authority, six guardrails, investigator-first.

July 20, 2026Read
Game Development in the Fable 5 Era: The AI-Assisted Asset and Code Pipeline That Actually Ships
ai-architecture1 min read

Game Development in the Fable 5 Era: The AI-Assisted Asset and Code Pipeline That Actually Ships

A three-person team can now scope a fifteen-person game — if the build runs as a pipeline. AI-assisted game dev in the Fable 5 era: direction bibles, verification gates, playtests.

July 17, 2026Read
Mobile App Development in the AI Era: On-Device Agents, the New Stack, and What Ships in 2026
ai-architecture1 min read

Mobile App Development in the AI Era: On-Device Agents, the New Stack, and What Ships in 2026

The AI feature demos on the founder’s flagship — and 60% of users can’t run it. The 2026 mobile stack: on-device models, hybrid routing, agents, battery budgets, fallbacks.

July 17, 2026Read
The AI Code Verification Bottleneck: Architecture for Reviewing at Generation Speed
ai-architecture1 min read

The AI Code Verification Bottleneck: Architecture for Reviewing at Generation Speed

Your team merges AI code faster than anyone can review it — and trust breaks first. The verification pyramid, blast-radius tiers, and org design for reviewing at generation speed.

July 16, 2026Read
WebGPU in Production 2026: The Browser Is Now a GPU Target — Graphics, Compute, and AI Inference
ai-architecture1 min read

WebGPU in Production 2026: The Browser Is Now a GPU Target — Graphics, Compute, and AI Inference

WebGPU hit baseline: compute shaders and AI inference in the browser, 10-15x on compute work. The production architecture — engine vs raw API, fallback ladders, in-page ML.

July 16, 2026Read
Prompt-Driven Graphics in 2026: Generating 3D and 2D Scenes with AI Without Shipping the Demo
ai-architecture1 min read

Prompt-Driven Graphics in 2026: Generating 3D and 2D Scenes with AI Without Shipping the Demo

Prompt a running 3D scene into existence in minutes — then watch it leak memory in production. The generate-direct-engineer pipeline that makes AI graphics shippable in 2026.

July 15, 2026Read
PixiJS vs Three.js in 2026: Choosing Your Web Graphics Engine Before It Chooses Your Roadmap
ai-architecture1 min read

PixiJS vs Three.js in 2026: Choosing Your Web Graphics Engine Before It Chooses Your Roadmap

Neither engine is better — one is 3D, one is 2D, and picking the mismatch costs a rewrite. PixiJS vs Three.js in 2026: the decision table, roadmap tiebreakers, and WebGPU on both.

July 15, 2026Read
Small Language Models in Healthcare 2026: Private, On-Prem Clinical AI That Fits Inside the Hospital
ai-architecture1 min read

Small Language Models in Healthcare 2026: Private, On-Prem Clinical AI That Fits Inside the Hospital

A specialized 7B model on the hospital’s own GPU never ships PHI off-site — and now matches larger models on medical exams. The private, on-prem clinical-AI architecture for 2026.

July 14, 2026Read
PixiJS in Production 2026: High-Performance 2D Web Graphics, WebGPU, and When 2D Beats 3D
ai-architecture1 min read

PixiJS in Production 2026: High-Performance 2D Web Graphics, WebGPU, and When 2D Beats 3D

The 2D particle scene flew on the laptop and melts on a real phone. PixiJS in production 2026: sprite batching, texture atlases, memory disposal, and WebGPU with WebGL fallback.

July 14, 2026Read
Three.js in Production 2026: WebGPU, Realtime, and the New Web-Graphics Standard
ai-architecture1 min read

Three.js in Production 2026: WebGPU, Realtime, and the New Web-Graphics Standard

The 3D prototype hit 120fps on the laptop and crawls on a real phone. Three.js in production 2026: draw calls, memory disposal, asset pipelines, and WebGPU with fallback.

July 14, 2026Read
AI SRE Agents: Architecture for Autonomous Incident Response
ai-architecture1 min read

AI SRE Agents: Architecture for Autonomous Incident Response

The same agent that mitigates an incident in 3 minutes can cause one in 3 seconds. AI SRE agents: the diagnose-and-remediate loop, the five guardrails, and where to start.

July 14, 2026Read
Why AI Proofs-of-Concept Die Before Production (and the Architecture That Ships)
ai-architecture1 min read

Why AI Proofs-of-Concept Die Before Production (and the Architecture That Ships)

The demo dazzled the board; six months later it still is not live. Why 62-70% of AI pilots die in the PoC-to-production gap, the five gaps, and the architecture that ships.

July 14, 2026Read
Embedding Models Comparison 2026: OpenAI vs Cohere vs Voyage vs BGE
ai-engineering1 min read

Embedding Models Comparison 2026: OpenAI vs Cohere vs Voyage vs BGE

Head-to-head comparison of the top embedding models in 2026: OpenAI text-embedding-3, Cohere Embed v3, Voyage AI, and BGE. Benchmarks, cost per 1M tokens, context windows, and a decision framework for RAG, code search, multilingual, and self-hosted deployments.

July 14, 2026Read
Embedding Models Comparison 2026: OpenAI vs Cohere vs Voyage vs BGE
ai-engineering1 min read

Embedding Models Comparison 2026: OpenAI vs Cohere vs Voyage vs BGE

Head-to-head comparison of the top embedding models in 2026: OpenAI text-embedding-3, Cohere Embed v3, Voyage AI, and BGE. Benchmarks, cost per 1M tokens, context windows, and a decision framework for RAG, code search, multilingual, and self-hosted deployments.

July 14, 2026Read
Embedding Models Comparison 2026: OpenAI vs Cohere vs Voyage vs BGE
ai-engineering1 min read

Embedding Models Comparison 2026: OpenAI vs Cohere vs Voyage vs BGE

Head-to-head comparison of the top embedding models in 2026: OpenAI text-embedding-3, Cohere Embed v3, Voyage AI, and BGE. Benchmarks, cost per 1M tokens, context windows, and a decision framework for RAG, code search, multilingual, and self-hosted deployments.

July 14, 2026Read
Embedding Models Comparison 2026: OpenAI vs Cohere vs Voyage vs BGE
ai-engineering1 min read

Embedding Models Comparison 2026: OpenAI vs Cohere vs Voyage vs BGE

Head-to-head comparison of the top embedding models in 2026: OpenAI text-embedding-3, Cohere Embed v3, Voyage AI, and BGE. Benchmarks, cost per 1M tokens, context windows, and a decision framework for RAG, code search, multilingual, and self-hosted deployments.

July 14, 2026Read
KV-Cache Offloading: Serving 10x More Users by Not Recomputing
ai-architecture1 min read

KV-Cache Offloading: Serving 10x More Users by Not Recomputing

Your GPU re-prefills the same 15,000-token prompt ten thousand times a day. KV-cache offloading to DRAM and NVMe turns that recompute into a cheap fetch — 10x users.

July 13, 2026Read
Feature Flag Architecture: Rollouts, Kill Switches, and Flag Debt
ai-architecture1 min read

Feature Flag Architecture: Rollouts, Kill Switches, and Flag Debt

Ship code to production without releasing it to everyone. Feature flags done right: release vs kill-switch types, sticky rollouts, failure defaults, and killing flag debt.

July 13, 2026Read
Distributed Locks: Redlock, Fencing Tokens, and Why Your Lock Doesn't Lock
ai-architecture1 min read

Distributed Locks: Redlock, Fencing Tokens, and Why Your Lock Doesn't Lock

A billing system double-charged 14 customers through a lock that worked exactly as documented. Redlock, fencing tokens, advisory locks, and when to delete the lock.

July 13, 2026Read
04Contact

Bring this thinkingto your business.

Start a
conversation

One essay, most weeks. No noise.