Guides
Guides
Direct answerFind reviewed guidance by task, from first requests and model selection to production troubleshooting.
Updated
AI API production reliability: timeouts, retries, limits, and SLOsDesign end-to-end deadlines, bounded retries, concurrency protection, circuit breaking, degradation, observability, and load acceptance for AI APIs.AI API quickstart: from API key to first requestLearn base URLs, API keys, model IDs, and endpoints, then make a first AI API call with curl and read the response.AI API security, privacy, and data governanceCreate verifiable boundaries for API keys, input data, file URLs, prompt injection, tool authorization, logs, and upstream retention.AI gateway admin runbook: channel rollout, retries, and supplier healthConfigure, test, canary, and roll back channels with least privilege, then diagnose retry, cooldown, catalog, account-balance, and supplier-health failures.AI model API platform beginner guide: from signup to calls and billing checksA first-time user's path through an AI model API aggregation platform, from account access and model selection to funding, API-key creation, first calls, billing checks, and safe support requests.API endpoint and capability matrixMap public gateway paths to capabilities, methods, transports, implementation status, and model-support boundaries without treating route existence as availability.Billing, usage, and cost reconciliationUnderstand runtime prices, groups, pre-consumption, final settlement, refunds, caching, and async tasks, then reconcile request and business cost.Choosing and calling coding modelsFilter coding models by repository size, tool use, reasoning depth, and latency, then validate on real tasks.Common AI API error codes and troubleshootingCompare official provider guidance for 400, 401, 402, 403, 404, 408, 409, 413, 422, 424, 429, 499, 5xx, and 529, including retryability, request IDs, and streaming failures.File and media input, output, and lifecycleChoose URL, Base64, data URL, or multipart safely, understand the unsupported Files boundary, and manage temporary media and result URLs.General text, chat, and reasoning model guideChoose general text models by reasoning depth, context, structured output, tools, latency, and cost.Image generation and vision model guideUse multimodal chat for image understanding and dedicated Images endpoints for generation and editing.Media and object-storage admin runbook: profiles, access, and lifecycleConfigure storage profiles for generated, staging, compliance, and message assets; validate S3 or OSS connectivity, private delivery, lifecycle, cleanup, and recovery.Model evaluation, canarying, and version migrationBuild a representative evaluation set and pinned baseline, then compare offline, canary, roll back, and monitor before a model upgrade or retirement.Model originators, hosted clouds, aggregators, and local runtimesSeparate model knowledge from the route used to call it across originators, hosted clouds, aggregators, workflows, local runtimes, and account bridges.RAG, embedding, and reranking model guideEvaluate chunking, embedding, retrieval, reranking, and evidence-grounded generation separately.SSE streaming and WebSocket realtime connectionsHandle SSE events, in-stream errors, cancellation, proxy buffering, backpressure, and realtime WebSocket sessions from first byte to confirmed completion.Speech, transcription, and realtime audio guideSpeech synthesis, transcription, and realtime conversation use distinct endpoints and transports.Tool use and AI agent model guideAgents should prioritize tool schemas, long-horizon state, failure recovery, and least privilege.Unified async tasks: submit, poll, and retain resultsUse the generic async wrapper and media task APIs while handling queues, progress, terminal states, JSON or binary results, expiration, and duplicate submission.Video generation and asynchronous API guideVideo generation is usually asynchronous: persist task IDs, limit polling, and validate result media.
兔子API