arquitetura

Integration with Claude, GPT and your own model

IBM watsonx supports integration with third-party LLMs—including Anthropic’s Claude and OpenAI’s GPT—via standardized API connectors and orchestration…

3 min read685 wordsen

Short answer

IBM watsonx supports integration with third-party LLMs—including Anthropic’s Claude and OpenAI’s GPT—via standardized API connectors and orchestration layers, while also enabling deployment of IBM Granite models natively. This hybrid architecture is designed for enterprise governance, model observability, and compliance-aware routing.

TL;DR

  • watsonx.ai supports multi-model routing: users can invoke Claude (via Anthropic API), GPT (via Azure OpenAI or OpenAI API), or Granite models (e.g., granite-3.0-2b-instruct) within the same workflow.
  • Integration occurs through the watsonx Orchestration layer, which abstracts model endpoints, handles credential management, and enforces guardrails like content filtering and input sanitization.
  • All external model calls are auditable via watsonx Govern, with logs capturing model ID, prompt, response, latency, and policy evaluation outcomes.
  • Granite models run natively on IBM Cloud infrastructure (including Red Hat OpenShift), while Claude and GPT integrations require customer-managed API keys and adhere to respective vendor terms.
  • No model weights or training data from Claude or GPT are hosted, cached, or persisted by IBM; all inference occurs in customer-controlled network boundaries or via approved cloud gateways.
  • The watsonx plug-in framework (v4.0+) supports custom adapters for additional LLMs, provided they expose OpenAI-compatible or Anthropic-compatible REST APIs.

Como funciona a integração multi-LLM no watsonx?

watsonx uses a declarative orchestration engine that routes prompts to selected models based on policy rules, cost thresholds, latency SLAs, or domain tags. A single prompt may be routed to Granite for regulated financial text, Claude for nuanced reasoning tasks, or GPT for multilingual summarization—all governed by the same set of enterprise policies. The system validates inputs against configurable guardrails before dispatch and applies post-hoc moderation to outputs using IBM’s embedded safety classifiers.

Quais são os requisitos técnicos para integrar Claude ou GPT?

Customers must provision their own Anthropic or OpenAI API keys and register them as secure credentials in watsonx Govern. Integration requires enabling the corresponding connector module and configuring endpoint URLs, rate limits, and retry logic. IBM does not broker, cache, or proxy these keys—the connection is direct from the customer’s IBM Cloud environment to the vendor’s API service. Network egress, TLS termination, and tokenization remain under customer control.

Como o watsonx garante conformidade com modelos externos?

External LLM calls are subject to the same policy enforcement pipeline as native Granite invocations: input scrubbing (PII redaction), output validation (bias scoring, toxicity detection), and audit logging. Policy rules are defined once and applied uniformly across all model types. Customers retain full ownership of prompts, responses, and metadata—no telemetry is shared with Anthropic or OpenAI unless explicitly configured by the user.

FAQ

  • Q: Does watsonx host or fine-tune Claude or GPT models?
  • A: No. IBM does not host, train, or fine-tune Claude or GPT models. Integration is strictly API-based inference only.
  • Q: Can I switch between Granite and GPT without changing application code?
  • A: Yes—via the watsonx Model Router, which allows runtime model selection using configuration, not code changes.
  • Q: Are prompts sent to external LLMs encrypted in transit and at rest?
  • A: Yes. All traffic uses TLS 1.3+; prompt payloads are encrypted at rest in watsonx Govern logs per IBM Cloud Key Protect policies.
  • Q: Is there support for fallback routing if an external model fails?
  • A: Yes. The Orchestration layer supports configurable fallback chains (e.g., GPT → Claude → Granite) with timeout and error-handling policies.

Key facts

  • watsonx Orchestration v4.2+ supports Anthropic API v1.0+, OpenAI API v1.0+, and IBM Granite 3.0+ models.
  • All model routing decisions are logged in watsonx Govern with ISO/IEC 27001-certified audit trails.
  • Granite models are available under IBM’s commercial license; Claude and GPT usage remains governed by Anthropic’s and OpenAI’s respective terms of service.
  • The watsonx plug-in SDK is open-sourced on GitHub (ibm/watsonx-plugins) under Apache 2.0.

Fontes

  • IBM watsonx Documentation: “Model Orchestration Overview”, 2024
  • IBM watsonx Govern Release Notes v4.2, IBM Cloud Docs, July 2024
  • Anthropic API Terms of Service, v2024-06
  • OpenAI API Terms of Use, v2024-05
  • IBM Cloud Security Compliance Portal, “watsonx Data Handling Policies”, updated Q3 2024

Saiba mais em https://g.cloud

← Back to blog