Best LLMs for Social & Promotional Content
Conciseness, platform-native conventions, engagement under tight character limits.
Conciseness, platform-native conventions, engagement under tight character limits.
Task-by-task breakdown
Public Response Generation
Drafts a concise public response to an input message using a supplied angle, evidence boundaries, channel profile, policy profile, and operator instructions. Illustrative uses include drafting a …
| Model | Quality (% of best) | Confidence | Overpay |
|---|---|---|---|
| GPT-5.6 Luna ★ | 95% | RANKED | best value |
| Tencent Hy3 | 92% | HIGH | 4.3x |
| GPT-5.6 Terra | 97% | HIGH | 7.9x |
| Qwen 3.7 Plus | 90% | HIGH | 13x |
| GPT-5.6 Sol | 100% | RANKED | 14x |
| Thinking Machines Inkling Small | 96% | HIGH | 16x |
| Gemini 3.5 Flash | 91% | HIGH | 17x |
| Thinking Machines Inkling | 93% | HIGH | 42x |
| Moonshot Kimi K3 best | 100% | RANKED | 100x |
Short Promotional Teaser Generation
Creates a grounded, editorial-style teaser for a content item from its name, title, source material, and optional source URL. Intended for compact discovery surfaces; returns 1–2 sentences in …
| Model | Quality (% of best) | Confidence | Overpay |
|---|---|---|---|
| GPT-5.6 Luna ★ | 95% | RANKED | best value |
| Tencent Hy3 | 99% | RANKED | 3.2x |
| NVIDIA Nemotron 3.5 Lightning | 92% | HIGH | 4.1x |
| GLM-5.3 Flash | 99% | RANKED | 6.7x |
| GPT-5.6 Terra | 95% | RANKED | 7x |
| DeepSeek V4 Flash | 95% | HIGH | 7.1x |
| NVIDIA Nemotron-3 Super 120B | 93% | RANKED | 9.7x |
| GPT-5.6 Sol | 94% | RANKED | 12x |
| NVIDIA Nemotron-3 Ultra 550B | 98% | RANKED | 17x |
| DeepSeek V4 Pro | 98% | HIGH | 20x |
| Gemini 3.5 Flash | 91% | RANKED | 24x |
| Qwen 3.7 Plus | 95% | RANKED | 25x |
| Claude Sonnet 5 best | 100% | RANKED | 28x |
| Meta Muse Spark 1.3 | 92% | MEDIUM | 32x |
| Thinking Machines Inkling Small | 97% | RANKED | 36x |
| Claude Opus 5 | 98% | HIGH | 38x |
| GLM-5.3 | 99% | RANKED | 65x |
| Moonshot Kimi K3 | 100% | RANKED | 85x |
| Tencent Hy4 Preview | 90% | MEDIUM | 92x |
| Grok 4.6 | 94% | HIGH | 106x |
| Thinking Machines Inkling | 96% | RANKED | 112x |
Promotional Campaign Generation
Generates a coordinated set of pre- and post-release promotional messages for a content item using supplied source material, campaign phases, channel rules, timing, message types, and count. …
| Model | Quality (% of best) | Confidence | Overpay |
|---|---|---|---|
| Gemini 3.5 Flash ★ best | 100% | RANKED | best value |
Promotional Message Generation
Generates one grounded promotional message for a content item using caller-supplied channel, audience, voice, length, link, and formatting rules. Illustrative uses include promoting a software …
| Model | Quality (% of best) | Confidence | Overpay |
|---|---|---|---|
| GPT-5.6 Luna ★ | 97% | RANKED | best value |
| GLM-5.3 Flash | 98% | MEDIUM | 3.5x |
| NVIDIA Nemotron-3 Super 120B | 91% | HIGH | 4x |
| GPT-5.6 Terra | 96% | RANKED | 4.5x |
| Gemini 3.8 Flash | 93% | RANKED | 4.9x |
| DeepSeek V4 Flash | 94% | RANKED | 7.6x |
| Thinking Machines Inkling Small | 94% | MEDIUM | 7.9x |
| GPT-5.6 Sol | 98% | RANKED | 9.7x |
| Gemini 3.5 Flash best | 100% | RANKED | 12x |
| Meta Muse Spark 1.3 | 93% | MEDIUM | 16x |
| Claude Opus 5 | 99% | MEDIUM | 19x |
| Thinking Machines Inkling | 97% | RANKED | 21x |
| DeepSeek V4 Pro | 95% | HIGH | 26x |
| Moonshot Kimi K3 | 98% | RANKED | 48x |
| Tencent Hy4 Preview | 96% | HIGH | 53x |
| GLM-5.3 | 98% | HIGH | 68x |
| Grok 4.6 | 98% | HIGH | 75x |
Community Content Promotion Generation
Decides whether a content item is appropriate for each supplied community and, for suitable communities, drafts a substantive contribution using caller-supplied channel and community rules. …
| Model | Quality (% of best) | Confidence | Overpay |
|---|---|---|---|
| MiniMax M3 ★ | 90% | RANKED | best value |
| GLM-5.3 Flash | 96% | MEDIUM | 1.5x |
| Qwen 3.7 Plus | 92% | HIGH | 2.6x |
| Thinking Machines Inkling Small | 91% | MEDIUM | 4.1x |
| Gemini 3.5 Flash best | 100% | RANKED | 6.6x |
| Thinking Machines Inkling | 90% | MEDIUM | 7.9x |
| Tencent Hy4 Preview | 91% | MEDIUM | 10x |
| DeepSeek V4 Pro | 93% | MEDIUM | 12x |
| Grok 4.6 | 96% | MEDIUM | 15x |
| GLM-5.3 | 97% | MEDIUM | 20x |
| Claude Sonnet 5 | 95% | RANKED | 21x |
| Moonshot Kimi K3 | 95% | HIGH | 32x |
Confidence — how sure we are about the quality score (more judgments + more agreement = higher confidence): RANKED many independent judges scored this model's outputs and their agreement is very high (most confident) — HIGH many judges have scored it and they mostly agree (well-pinned) — MEDIUM enough judges have weighed in to publish, but they disagree more than we'd like (treat with a small grain of salt). LOW-confidence cells are hidden everywhere on the site. See the methodology for the exact thresholds.