Skip to content
ai-platform-engineeravailable for select projects

I build the platform under the product.

I'm Ravi Roy — an AI Platform Engineer & Infrastructure Architect. I design and ship production AI platforms, multi-tenant SaaS, agents, and the orchestration layer behind them — end to end, from Postgres to LLM routing to mobile.

  • AI Platforms
  • AI Infrastructure
  • Workflow Automation
  • Multi-Tenant SaaS
  • Voice AI
  • Mobile & Web
60+
AI providers routed
10+
Products shipped
7+ yrs
Full-stack
Multi-tenant
By default
what-i-build

Not websites. Platforms, agents, and infrastructure.

Most of what I build is the hard layer underneath AI products — the orchestration, identity, and multi-tenancy that make them work at scale.

AI Platforms & Orchestration

The runtime that makes AI products possible — routing, fallback, and multi-model architecture across 60+ providers.

  • AI orchestration backend (Mastra)
  • Provider-agnostic model routing
  • Multi-provider architecture
  • Organization-aware AI memory

AI Agents & Voice AI

Agents that use tools and remember context, and voice systems that hold real conversations.

  • Tool-calling agent runtimes
  • Low-latency voice pipelines (STT → LLM → TTS)
  • Interruption / barge-in handling
  • RAG & tenant-scoped memory

Workflow Automation

Node-based pipelines and content distribution that run reliably, async, at scale.

  • Node-based workflow engine
  • Scheduling & async execution
  • Retries & queue processing
  • Multi-channel content distribution

Auth, Identity & Multi-Tenancy

The hard identity layer — solved once, reused everywhere.

  • Multi-tenant orgs & projects
  • RBAC, SSO, passwordless
  • Audit logs & session management
  • Encrypted secret vault

Mobile & Web Products

End-to-end product delivery, from data model to app store.

  • React Native / Expo apps
  • Next.js / React / Vue / Ionic web
  • Production backends (Fastify / Express / NestJS)
  • Hardware / IoT integration
featured-projects

Flagship work.

The platforms and products that define what I build — each a real, live system. Real screenshots, real brand, real scale.

01
Elevence AI logoAI Platform

Elevence AI

One studio for GPT-5, Claude, Gemini & 60+ models — unified behind a single interface.

Next.jsTypeScriptNode.jsPostgreSQLOpenAI
02
Voice AI

VoAgents AI

AI voice agents & phone assistants that answer, schedule, and convert calls automatically.

TypeScriptNode.jsMastraOpenAIWebSocketsTelephony
03
ClipCam logoConsumer App

ClipCam

Reaction clips in seconds — record over any clip, auto-stitched in the cloud.

React NativeExpoNode.jsTypeScriptFFmpegCloud Render
04
EXL AI Playground logoAI Playground

EXL AI Playground

An interactive playground to experiment with language models, image generation, and AI tools — in real time.

Next.jsTypeScriptNode.jsOpenAI
05
Fanisin logoConsumer AI

Fanisin

Where every fan is in — live video, voice, and text with AI twins of creators.

Next.jsNode.jsTypeScriptPostgreSQLOpenAIWebRTC
06
HiHelloHR logoEnterprise / HR

HiHelloHR

A full HRMS — payroll, attendance, facial recognition, and IoT access control.

React NativeNode.jsTypeScriptMongoDBIoTComputer Vision
ai-infrastructure

The orchestration layer behind the products.

The single most useful thing I build: the infrastructure that makes AI products possible without rebuilding the foundation every time.

60+
model providers routed through one API
1
provider-agnostic orchestration layer
fallback paths — never locked to one vendor
01

Model routing

Provider-agnostic router with cost-aware selection and automatic fallback across 60+ providers.

02

Multi-model architecture

A single request can fan out or fail over between models — no app changes required.

03

Organization-aware memory

Agent context isolated per tenant. Memory never leaks across organizations.

04

Workflow automation

An orchestration backend that chains models, tools, and steps into reliable workflows.

infrastructure-and-automation

The platforms underneath the products.

Internal infrastructure I've built and run in production — the hard layer other products delegate to. Solved once, correctly, and reused everywhere.

01One API in front of 60+ model providers.

AI Orchestration Platform

A provider-agnostic orchestration layer that routes every AI request across 60+ providers with fallback and cost-aware selection. Apps ask for a capability; the platform decides which model answers — so products are never locked to a single vendor.

  • 60+ AI providers behind one interface
  • Model routing with fallback & cost-aware selection
  • Multi-provider architecture (fan-out / fail-over)
  • AI workflows chaining models + tools
  • Mastra backend (Bun / Hono) runtime
  • Organization-aware AI memory
Mastra·Bun·Hono·TypeScript·PostgreSQL·Redis·Pinecone
02Node-based pipelines that run reliably, async, at scale.

Workflow Orchestration Engine

A node-based workflow engine for automating multi-step pipelines — scheduled or event-driven, executed asynchronously with retries and backoff. The substrate behind automations across the platform.

  • Node-based visual workflows
  • Scheduling (cron + event triggers)
  • Asynchronous execution
  • Automatic retries with backoff
  • Multi-step pipelines with branching
  • Durable queue processing
Node.js·TypeScript·Redis·BullMQ·PostgreSQL
03Publish once, distribute everywhere — on schedule.

Content Distribution Infrastructure

A multi-channel distribution system that schedules, repurposes, and auto-publishes content across social platforms and blogs through OAuth integrations and queue-backed processing.

  • Scheduling & auto-publishing
  • Content repurposing per channel
  • Multi-channel distribution
  • OAuth integrations
  • Queue processing
Channels
LinkedInX / TwitterFacebookInstagramWordPressGhostDev.toCustom Blogs
Node.js·TypeScript·Redis·OAuth 2.0·PostgreSQL
04Multi-tenant identity, solved once for every product.

Authentication Platform

The org-first identity platform every product delegates to: multi-tenant organizations and projects, RBAC, SSO, passwordless and social login, session management, audit logs, and an encrypted secret vault.

  • Multi-tenant — organizations & projects
  • RBAC (org / role / resource)
  • SSO + passwordless + email OTP
  • Social login
  • Session management
  • Audit logs
  • Encrypted secret vault
NestJS·Node.js·TypeScript·PostgreSQL·Redis·JWT
more-products

More products.

Consumer apps, AI tools, and platforms — built on the same orchestration and identity foundation, shipped end to end.

All work →
technical-expertise

A modern, full-stack toolkit.

From the AI layer down to the database and out to mobile — the stack I use to ship platforms end to end.

AI & Infrastructure
LLM InfrastructureAI OrchestrationModel RoutingAI AgentsVoice AIRAG / MemoryLangChainMastraOpenAIGoogle GeminiOllamaTensorFlow
Backend
Node.jsNestJSFastifyExpressHonoBunTypeScriptGraphQLgRPCKeycloak
Data & Vector
PostgreSQLMongoDBMySQLRedisPineconeChromaFirebaseDrizzleMongooseSequelize
Frontend & Mobile
ReactNext.jsReact NativeExpoVue.jsIonicCordovaThree.jsTailwind CSSshadcn/ui
Cloud & DevOps
AWSGoogle CloudDockerKubernetesJenkinsVercelRailwayHerokuCI/CDFirebase
Architecture
Multi-Tenant SaaSRBAC / SSOWorkflow AutomationSystem DesignSecret VaultsIoT Integration
writing

From the notebook.

Deep-dives on AI infrastructure, orchestration, and the systems behind the products — newest first.

All writing →
Emerging Technologies: Specialized Hardware & Edge AI Architectures
InsightsSep 18, 2026 · 19 min read

Emerging Technologies: Specialized Hardware & Edge AI Architectures

Emerging Technologies for edge AI explained with key architectures, tools, and benefits. Learn what to watch next and plan smarter.

Architecting Production-Ready Mobile App Development with Cloud AI Integration
InsightsSep 18, 2026 · 16 min read

Architecting Production-Ready Mobile App Development with Cloud AI Integration

Master robust mobile app development for AI with proven cloud integration patterns. Ensure your AI apps are production-ready. Learn more!

Optimizing React Native for On-Device AI: A Guide to Mobile App Development
InsightsSep 17, 2026 · 15 min read

Optimizing React Native for On-Device AI: A Guide to Mobile App Development

Mobile App Development strategies for on-device AI inference—boost speed, privacy, and offline performance. Learn how to optimize now.

Evaluating Open-Source LLMs for Production Generative AI Apps
InsightsSep 17, 2026 · 16 min read

Evaluating Open-Source LLMs for Production Generative AI Apps

Unlock the full potential of Generative AI. Learn to evaluate & benchmark open-source LLMs for robust production applications. Start building better AI today!

Securing SaaS Products: Data Isolation & Access Control Best Practices
InsightsSep 16, 2026 · 15 min read

Securing SaaS Products: Data Isolation & Access Control Best Practices

SaaS Products security best practices to protect tenant data, tighten access control, and reduce risk. Read the full guide.

Building Custom Voice AI: Open-Source LLM, ASR, & TTS Guide
InsightsSep 16, 2026 · 16 min read

Building Custom Voice AI: Open-Source LLM, ASR, & TTS Guide

Unlock the potential of Voice AI to create personalized assistants. Learn how open-source LLMs, ASR, and TTS empower custom solutions. Start building!

Implementing Distributed Caching Strategies for High-Performance Software Engineering
InsightsSep 15, 2026 · 21 min read

Implementing Distributed Caching Strategies for High-Performance Software Engineering

Software Engineering teams can speed up backend performance with proven distributed caching strategies. Improve latency and scale—start optimizing now.

Production AI Innovation: Multimodal Full-Stack Application Patterns
InsightsSep 15, 2026 · 14 min read

Production AI Innovation: Multimodal Full-Stack Application Patterns

AI Innovation strategies for production full-stack apps—learn practical multimodal patterns, reduce complexity, and move faster with confidence.

Full-stack Observability for AI: From UX to LLM Latency
InsightsSep 14, 2026 · 14 min read

Full-stack Observability for AI: From UX to LLM Latency

Unlock peak AI performance with full-stack observability. Monitor frontend UX, LLM latency & more to optimize your AI applications. Learn how!

Advanced Generative AI Strategies for Production Deployment
InsightsSep 14, 2026 · 15 min read

Advanced Generative AI Strategies for Production Deployment

Generative AI strategies for production teams—improve reliability, quality, and scale with practical prompt engineering. Read more.

contact

Building something that needs real infrastructure?

If you're building an AI platform, a multi-tenant SaaS, or the systems underneath one — let's talk. I take on a small number of serious projects.

available for select projects