000%
Ali Saad / Creative systems practiceSignal acquired
DestinationDestination
Destination / Destination11 / route handoff

Back to the archiveW—05 / 07
05
LLM-Hub Pro Max neon provider routing artwork
Ali Saad (Self-initiated product)2026

Selected work / Full-Stack Developer & System Architect

LLM-HubProMax

Full-Stack DevelopmentAPI ArchitectureAI InfrastructureSelf-Hosted Deployment
Scroll to enter
01The project, in one breath

The core challenge was normalizing providers with different model catalogs, capabilities, rate limits, error behavior, and endpoint coverage behind one predictable OpenAI-style interface. The gateway uses a provider catalog, capability metadata, explicit routing headers, sticky…

Self-hosted OpenAI-compatible API gateway and admin dashboard that routes applications across multiple AI providers through one /v1 interface, with sticky fallbacks, health checks, encrypted local keys, analytics, and multimodal testing.

Role
Full-Stack Developer & System Architect
Client
Ali Saad (Self-initiated product)
Year
2026
02The build, reduced to what mattered

Three decisions shaped the final experience.

03
01
ONE CLIENT INTERFACE

Unified OpenAI-Compatible /v1 API

One proxy surface covers model listing, streaming and non-streaming chat, tool calling, embeddings, image generation and edits, speech, transcription, translation, and realtime…

02
ROUTING CONTROL

Sticky Fallback Routing

Configurable chains keep requests on the selected route when healthy, apply cooldowns after rate limits or provider errors, and expose routed-provider and fallback-attempt metadata in…

03
LIVE AVAILABILITY

Provider Health & Model Discovery

Health checks, model availability discovery, capability tracking, provider ranking, and model sweeps show which routes are configured and currently usable across the provider catalog.

OutcomeWhat remained after the build

LLM-Hub Pro Max provides one working `/v1` entry point for applications that need to move between models and providers without rewriting client integrations. It supports routed chat, embeddings, images, audio, and realtime session-token workflows, with provider and fallback metadata returned alongside requests. The dashboard centralizes key management, fallback order, model availability, health checks, logs, analytics, diagnostics, and multimodal testing. When a route fails or rate-limits, cooldown and fallback…

OpenAI-compatible /v1 routesAPI Surface
Self-hosted and local-firstDeployment
Personal and small-team AI workloadsTarget
Next projectW—06 / 07
Continue the archive

Portfolio Control Room

Enter project
Portfolio Control Room
Have something worth building?Start a project
aliihsaad1@gmail.com
InstagramLinkedInGitHubFacebook
© 2026 Ali Saad