Kimi-K3
Kimi K3 is an open-weight, 2.8T-parameter native multimodal agentic model with a 1M-token context window, designed for frontier coding, knowledge work, and reasoning tasks.
It ships as a standalone Go runtime that embeds FrankenPHP, Mercure, and a full Laravel app, making self-hosting as simple as downloading a binary or running a Docker container.
English | 简体中文
HelmDesk is an open-source, self-hosted, AI-native customer support system for small and medium-sized teams. It brings omnichannel customer conversations into one workspace for AI-powered support and team collaboration.
With a translation provider configured, HelmDesk can translate incoming and outgoing messages into each teammate's preferred language. The inbox keeps the original text and translation together, preserving the conversation context when agents hide or refresh a translation.
Agents can write a reply in the language they know best, preview the translation the visitor will receive, and confirm it before sending.
The diagram below shows initial routing, AI and human handoffs, timeout-driven transitions, closing, and reopening.
PHP 8.5 ZTS, Composer, Node.js, Go, and php-config are required.
git clone git@github.com:helmdesk-ai/helmdesk.git
cd helmdesk
composer setup
make
Once started, open http://localhost:8080.
Build artifacts are stored in build/output.
Every build requires a SemVer such as 1.0.0. Make commands use APP_VERSION, and Po
AI-native, self-hosted customer support for small and medium-sized teams. 面向中小团队的 AI 原生自托管客服系统。
Similar projects matched by category, topics, and programming language.
Kimi K3 is an open-weight, 2.8T-parameter native multimodal agentic model with a 1M-token context window, designed for frontier coding, knowledge work, and reasoning tasks.
Harness Engineering is a methodology for improving coding agent outputs by carefully crafting the environment around them—providing curated context, tools, and executable constraints that encode an organization’s nonfunctional requirements and cumulative lessons.
A 28.9 million parameter language model runs on an $8 ESP32-S3 microcontroller entirely on-device, generating simple stories at about 9.5 tokens per second.