OpenAI-compatible text embeddings from a private, self-hosted local encoder (768 dims). Pay per call in USDC on Base via x402 - no API key or account. Send {"input": string or array of up to 64 texts}; receive standard OpenAI embeddings JSON with usage. Built for RAG, semantic search, dedupe and clustering in agent pipelines. ~2,000-token limit per text. Same privacy-first operator as our Private LLM Inference chat tiers.