Skip to main content
Telara

Integrations / NLP Cloud (direct)

NLP Cloud (direct) logo

NLP Cloud (direct)

Direct NLP Cloud API calls (nlpcloud SDK or raw HTTP, no framework). NLP Cloud is a multi-provider aggregator: one API surfaces a catalog of open-source and proprietary text-generation/embedding models behind a single key, at https://api.nlpcloud.io/v1/<model>/<endpoint> (well- documented publicly; NLP Cloud's own docs pages returned 403 to programmatic fetch during this session, so the base URL and per-model path shape could not be independently re-verified live here — treat the JSONPaths below as the standard gen_ai.* convention, not a vendor-confirmed field layout). Unlike Voyage/Jina/Mixedbread, NLP Cloud fronts chat/completion-style models as well as embeddings, so an output_tokens metric is meaningful here and is kept. IMPORTANT for chargeback: because NLP Cloud is an aggregator, the OTLP spans this connector ingests only carry NLP Cloud's own request/response — whichever underlying model actually served the call, cost and usage land on NLP Cloud as the billing entity, not on the underlying model vendor. Per- underlying-model attribution is only possible if the customer's own instrumentation additionally captures the model id NLP Cloud reports back (gen_ai.response.model), which this connector does try via model_fallback_paths. Auth is NOT bearer-prefixed: NLP Cloud expects `Authorization: Token <api_key>` (not `Bearer <api_key>`), reflected below via header_prefix. NLP Cloud has no admin/billing API of its own beyond the dashboard, so spend is only capturable via this customer-side OTLP push. No telara-observability provider wrapper exists for NLP Cloud today (the package's optional-dependencies only cover langchain/langgraph/ crewai/autogen/llamaindex/openai-agents/openai/anthropic/azure-openai) — spans must come from the customer's own OTel instrumentation around the NLP Cloud SDK/HTTP call. Spans should be tagged telara.span.kind=provider so the cost rollup excludes them when an active framework adapter already covers the same call (double-count guard).

OTLP
Spend

What Telara does

Capability matrix

Before you connect

Prerequisites

A process that uses NLP Cloud (direct) and can export OpenTelemetry.
Ability to run the Telara CLI where that process runs.

Vendor setup

Get the credential

1
Print the OTLP endpoint
Run `telara otlp`. It prints the endpoint and headers Telara expects. Do not paste a vendor API key on this connector.
2
Point NLP Cloud (direct) at it
Enable OpenTelemetry in NLP Cloud (direct) (or in your collector) and send traces to the endpoint that command prints.

In Telara

Connect in Telara

OTLP

This connector does not take a vendor key in Settings → Integrations. Run `telara otlp` (or point the vendor’s OpenTelemetry exporter at the endpoint that command prints) so usage arrives as telemetry.

Telara has not published the exact form fields for this method yet. Complete whatever the Connect dialog shows.

Spend

What gets measured

Resolution: Per request.

Attribution: Matched to people by email.

How far back vendor history goes is not published uniformly for this connector.

  • input_tokens
    tokens
  • output_tokens
    tokens
  • framework_runs
    count

Sync

Freshness & sync

Spend connectors refresh on the vendor pull cadence declared in the catalog. Indexing lag is not published for this connector.

Data handling

Permissions & data handling

Telara has not published a permission list for this connector yet. Treat the vendor consent screen as the source of truth until this page lists scopes.

This catalog entry does not declare extra semantic-search exclusions.

After connect

Verify it worked

This connector does not declare a credential probe yet. A saved connection means Telara stored the credential, not that the vendor accepted it.

Honesty

Known limitations

Nothing appears in Telara until that process actually emits telemetry.
This is not an admin-key connector. Org-level spend, if Telara offers it for this vendor, is a separate page.