Tina4

AI Client#

One class. Three methods. No provider SDK. The Tina4 AI client sends chat, completion, embedding, and streaming requests through one provider-neutral API.

This chapter covers the application-facing Ai class. The tina4 ai command serves a different purpose: it installs Tina4 skills and context for coding assistants.

Configure a provider#

The client supports local, openai, and anthropic. Local mode targets an OpenAI-compatible endpoint and needs no API key.

ini
TINA4_AI_PROVIDER=localTINA4_AI_URL=http://localhost:11437TINA4_AI_MODEL=llama3.2TINA4_AI_TIMEOUT=60TINA4_AI_CONNECT_TIMEOUT=10TINA4_AI_MAX_RETRIES=2

Use the hosted OpenAI service by changing three values:

ini
TINA4_AI_PROVIDER=openaiTINA4_AI_URL=https://api.openai.com/v1TINA4_AI_MODEL=gpt-4o-miniTINA4_AI_KEY=your-api-key

Anthropic uses the same keys:

ini
TINA4_AI_PROVIDER=anthropicTINA4_AI_URL=https://api.anthropic.com/v1TINA4_AI_MODEL=claude-3-5-haiku-latestTINA4_AI_KEY=your-api-key

An explicit method argument wins over the environment. The environment wins over the built-in default. This lets one request use another model without changing the rest of the process.

Complete one prompt#

Ai.complete sends one user message and returns the response text.

python
from tina4_python import Aiโ€‹answer = Ai.complete("Explain why this query needs an index")print(answer)

Pass request options when one call needs a different model or provider:

python
answer = Ai.complete(    "Summarize this incident report",    provider="openai",    model="gpt-4o-mini",    temperature=0.2,    max_tokens=300,    timeout=20,)

Hold a chat#

Ai.chat accepts a non-empty list of messages. Each message needs a system, user, or assistant role and string content.

python
from tina4_python import Aiโ€‹response = Ai.chat([    {"role": "system", "content": "Answer as a concise database engineer."},    {"role": "user", "content": "When should I use a composite index?"},])โ€‹print(response.text)print(response.model)print(response.usage["total_tokens"])print(response.finish_reason)

The ChatResponse object carries five fields:

FieldMeaning
textNormalized response text
modelModel reported by the provider
usageprompt_tokens, completion_tokens, and total_tokens
finish_reasonProvider finish reason, or None
rawOriginal provider response

Use raw only when you need provider-specific metadata. Keep application logic on the normalized fields so a provider change does not spread through your code.

Stream text#

Set stream=True to receive ordered text deltas. The iterator yields text only and ignores provider metadata events.

python
from tina4_python import Aiโ€‹chunks = Ai.chat(    [{"role": "user", "content": "Write a short release announcement."}],    stream=True,)โ€‹for chunk in chunks:    print(chunk, end="", flush=True)

Tina4 may retry before the first delta arrives. It never retries after yielding text because that could duplicate content the caller has already displayed.

Create embeddings#

Ai.embed preserves the input shape. One string returns one vector. A list returns one vector per input, in the same order.

python
from tina4_python import Aiโ€‹vector = Ai.embed("Tina4 keeps application code small")โ€‹vectors = Ai.embed([    "The router finds a matching handler",    "The ORM maps a row to a model",])

Set TINA4_EMBED_URL when the embedding service uses a different base URL. Anthropic does not expose embeddings through this contract, so provider="anthropic" raises AiConfigError.

Handle failures#

All client failures inherit from AiError:

ErrorMeaning
AiConfigErrorMissing key, invalid provider, bad option, or unsupported capability
AiHTTPErrorProvider returned a failing HTTP status or the transport failed
AiTimeoutErrorConnection or total request deadline expired
AiParseErrorA successful response did not match the provider contract
python
from tina4_python import Ai, AiErrorโ€‹try:    print(Ai.complete("Create a migration plan"))except AiError as error:    print(f"AI request failed: {error}")

Hosted providers fail before sending when TINA4_AI_KEY is missing. Error text never includes the key, prompt, or provider response body. A failure stays a failure; Tina4 never turns it into an empty answer.

Timeouts and retries#

TINA4_AI_CONNECT_TIMEOUT bounds connection setup. TINA4_AI_TIMEOUT bounds the whole request, including retries and response reads. TINA4_AI_MAX_RETRIES applies only to connection failures, HTTP 429, and HTTP 5xx responses.

Other HTTP 4xx responses and malformed successful responses run once and fail. The client carries transient trouble for a bounded distance, then hands the error back to your code.