AI Client#
One class. Three methods. No provider SDK. The Tina4 AI client sends chat, completion, embedding, and streaming requests through one provider-neutral API.
This chapter covers the application-facing Ai class. The tina4 ai command serves a different purpose: it installs Tina4 skills and context for coding assistants.
Configure a provider#
The client supports local, openai, and anthropic. Local mode targets an OpenAI-compatible endpoint and needs no API key.
TINA4_AI_PROVIDER=localTINA4_AI_URL=http://localhost:11437TINA4_AI_MODEL=llama3.2TINA4_AI_TIMEOUT=60TINA4_AI_CONNECT_TIMEOUT=10TINA4_AI_MAX_RETRIES=2Use the hosted OpenAI service by changing three values:
TINA4_AI_PROVIDER=openaiTINA4_AI_URL=https://api.openai.com/v1TINA4_AI_MODEL=gpt-4o-miniTINA4_AI_KEY=your-api-keyAnthropic uses the same keys:
TINA4_AI_PROVIDER=anthropicTINA4_AI_URL=https://api.anthropic.com/v1TINA4_AI_MODEL=claude-3-5-haiku-latestTINA4_AI_KEY=your-api-keyAn explicit method argument wins over the environment. The environment wins over the built-in default. This lets one request use another model without changing the rest of the process.
Complete one prompt#
Ai.complete sends one user message and returns the response text.
from tina4_python import Aiโanswer = Ai.complete("Explain why this query needs an index")print(answer)Pass request options when one call needs a different model or provider:
answer = Ai.complete( "Summarize this incident report", provider="openai", model="gpt-4o-mini", temperature=0.2, max_tokens=300, timeout=20,)Hold a chat#
Ai.chat accepts a non-empty list of messages. Each message needs a system, user, or assistant role and string content.
from tina4_python import Aiโresponse = Ai.chat([ {"role": "system", "content": "Answer as a concise database engineer."}, {"role": "user", "content": "When should I use a composite index?"},])โprint(response.text)print(response.model)print(response.usage["total_tokens"])print(response.finish_reason)The ChatResponse object carries five fields:
| Field | Meaning |
|---|---|
text | Normalized response text |
model | Model reported by the provider |
usage | prompt_tokens, completion_tokens, and total_tokens |
finish_reason | Provider finish reason, or None |
raw | Original provider response |
Use raw only when you need provider-specific metadata. Keep application logic on the normalized fields so a provider change does not spread through your code.
Stream text#
Set stream=True to receive ordered text deltas. The iterator yields text only and ignores provider metadata events.
from tina4_python import Aiโchunks = Ai.chat( [{"role": "user", "content": "Write a short release announcement."}], stream=True,)โfor chunk in chunks: print(chunk, end="", flush=True)Tina4 may retry before the first delta arrives. It never retries after yielding text because that could duplicate content the caller has already displayed.
Create embeddings#
Ai.embed preserves the input shape. One string returns one vector. A list returns one vector per input, in the same order.
from tina4_python import Aiโvector = Ai.embed("Tina4 keeps application code small")โvectors = Ai.embed([ "The router finds a matching handler", "The ORM maps a row to a model",])Set TINA4_EMBED_URL when the embedding service uses a different base URL. Anthropic does not expose embeddings through this contract, so provider="anthropic" raises AiConfigError.
Handle failures#
All client failures inherit from AiError:
| Error | Meaning |
|---|---|
AiConfigError | Missing key, invalid provider, bad option, or unsupported capability |
AiHTTPError | Provider returned a failing HTTP status or the transport failed |
AiTimeoutError | Connection or total request deadline expired |
AiParseError | A successful response did not match the provider contract |
from tina4_python import Ai, AiErrorโtry: print(Ai.complete("Create a migration plan"))except AiError as error: print(f"AI request failed: {error}")Hosted providers fail before sending when TINA4_AI_KEY is missing. Error text never includes the key, prompt, or provider response body. A failure stays a failure; Tina4 never turns it into an empty answer.
Timeouts and retries#
TINA4_AI_CONNECT_TIMEOUT bounds connection setup. TINA4_AI_TIMEOUT bounds the whole request, including retries and response reads. TINA4_AI_MAX_RETRIES applies only to connection failures, HTTP 429, and HTTP 5xx responses.
Other HTTP 4xx responses and malformed successful responses run once and fail. The client carries transient trouble for a bounded distance, then hands the error back to your code.