Hugging Face × WorkDaddy: LLM Provider Integration (2026)

Use Hugging Face as the engine behind your WorkDaddy agent team: 56 models available, context windows up to 1M tokens, connected in minutes with your own key.

Hugging Face is one of 178 LLM providers supported by WorkDaddy’s model layer in 2026, in the Inference Clouds group. The catalog currently lists 56 models from Hugging Face: 46 with explicit reasoning, 18 with vision input, and 54 with tool calling — the capability that matters most for agent work. The largest context window reaches 1M tokens.

At a glance

Models available 56
Reasoning models 46
Vision models 18
Tool-calling models 54
Largest context window 1M tokens
API endpoint https://router.huggingface.co/v1
API key variable HF_TOKEN
SDK adapter @ai-sdk/openai-compatible
Provider docs huggingface.co ↗

Why run WorkDaddy agents on Hugging Face in 2026

Tool calling, context size, and cost decide how well a provider drives an agent team. Hugging Face brings 54 tool-calling models to WorkDaddy — enough to power the full loop of storefront edits, analytics queries, and content production — with 1M tokens of context at the top end for whole-codebase and long-report work. Because WorkDaddy is model-agnostic, you can route heavy reasoning to Hugging Face’s strongest model and bulk work to its cheapest, inside one team.

Best Hugging Face models for e-commerce agents (2026)

The current flagships from Hugging Face in the WorkDaddy catalog, newest first:

  • Inkling Small — Efficient model for low-latency assistance, extraction, and routine automation (524K context)

  • Kimi K3 — Kimi multimodal agent model for visual understanding, coding, and planning (1M context)

  • Inkling — Multimodal model for analyzing text, images, documents, and rich media (1M context)

  • Hy3 — Tencent Hy reasoning model for coding, instruction following, and agent tasks (262K context)

  • GLM-5.2 — Open flagship GLM for long-horizon coding agents and million-token context work (262K context)

  • Kimi K2.7 Code — Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking (262K context)

How to connect Hugging Face to WorkDaddy

WorkDaddy talks to Hugging Face at router.huggingface.co via the @ai-sdk/openai-compatible adapter. Add your credential (typically HF_TOKEN) in Settings → Models, pick a default model, and every agent can use it immediately — or add Hugging Face as one engine among several and let tasks route to the best fit.

Featured models

Model Context Reasoning Released
Inkling Small 524K Yes 2026-07-30
Kimi K3 1M Yes 2026-07-16
Inkling 1M Yes 2026-07-15
Hy3 262K Yes 2026-07-06
GLM-5.2 262K Yes 2026-06-13
Kimi K2.7 Code 262K Yes 2026-06-12

How it works

01

Connect

In WorkDaddy, open Settings → Models, choose Hugging Face, and paste your API key (HF_TOKEN). No key? Start on the managed gateway instead.

02

Put agents to work

Pick Inkling Small or any of the 56 available models as your default — per-agent overrides let you match model to task.

03

Review and approve

Agent output stays draft-first regardless of the model: review diffs and approve actions exactly as before.

Frequently asked questions

Does WorkDaddy support Hugging Face in 2026?

Yes — Hugging Face is a supported LLM provider with 56 models in the WorkDaddy catalog, connected with your own API key via the @ai-sdk/openai-compatible adapter.

How many Hugging Face models can I use with WorkDaddy?

56 models are listed for Hugging Face, including 46 reasoning models and 18 vision-capable models. The flagship lineup above shows the newest.

What context window do Hugging Face models offer?

Up to 1M tokens on the largest model — relevant for whole-repo storefront work and long analytics reports.

Is Hugging Face the best LLM provider for e-commerce agents?

It depends on the task mix — Hugging Face sits in the Inference Clouds group. WorkDaddy is model-agnostic, so the practical answer in 2026 is to combine providers: route each agent task to whichever model fits best.

Put the team to work on your store

Connect your storefront, analytics, and email stack, set a goal, and let the agents run the work end to end. Start on the free plan with your own model key.