# Grok Build (xAI)

Feedback Bench, coding agents, built 2026-10-01, window 2026-08-31 to 2026-09-27. Web page: https://feedbackbench.com/#/agent/grok-build

| Measure | Value |
|---|---|
| Rank | 16 of 17 (rank range 16–16) |
| Feedback Score | 12.6 (95% interval 12.4–12.7) |
| Popularity | 0.031 (share of voice 0.07%) |
| Customer love | 0.517 (95% interval 0.506–0.527) |
| Top quadrant | no |
| Authors | 66 |
| Posts counted | 78 |
| Posts that judge the agent | 51 |
| Criteria better / worse than peers | 0 / 0 of 63 |

## The brief

Written by Claude Opus 5.5 from 24 labelled posts and the numbers on this page. Interpretation, not measurement: every quote is verbatim and links to its post.

**A rule-following workhorse on generous limits, shadowed by privacy doubts.**

TL;DR:

- Users praise generous plan limits and rule-following; several run it as a subagent under Claude or Codex.
- Privacy is the sharpest complaint, from a reported cross-tenant session leak to advice to route through Cursor.
- Brand confusion and thin editor integration push users toward Cursor for the IDE surface.

### What hurts

- **Privacy trust is the weak point** ([Data retention, training use and deployment isolation](https://feedbackbench.com/criteria/account.data_privacy.md)). Users doubt Grok Build's data handling, citing a reported session-isolation failure and recommending Cursor's wrapper as the safer path to the same model.
  Every data-privacy post is a complaint. One user describes an empty prompt returning tool calls against another user's workspace and argues that closing it as a hallucination proves nothing. That is a user report, not a confirmed finding.
  
  Others draw a practical line. Run Grok through Cursor, not the native CLI, and Cursor's privacy mechanisms apply. Another post names Grok Build as a weaker zero-retention bet than Anthropic.
  Evidence:
  - Complaint, Grok Build, r/codex, 2026-09-18: “i’ve done a bit of research. i think if you use grok through cursor, not grok cli, it uses cursor’s mechanisms of getting the agent to do the task. not grok’s. and cursor’s privacy is much much higher. if your argument is more about morality than your codebase’s privacy, that’s also completely valid.” [source](https://www.reddit.com/r/codex/comments/1wjhw4l/is_claude_a_value_switch_now/pamwwss/)
  - Complaint, Grok Build, r/AI_Agents, 2026-09-13: “**fyi: malicious actors could likely hijack your grok build sessions during the month of june by simply prompting 'hi'** this is serious because it is not a chatbot making up a story. a stateless "hi" with tools: \[\] still came back finish\_reason: tool\_calls and executed read\_file/grep on another user's workspace. that means session isolation failed at the serving layer: one tenant's context was reachable from another. if that happens, a prompt as empty as "hi" can pull someone else's files, tools, and private session. closing it as a hallucination, then deprecating the model, does not prove the mix-up cannot happen on whatever replaced it.” [source](https://www.reddit.com/r/AI_Agents/comments/1wexro6/fyi_malicious_actors_could_likely_hijack_your/)
  - Complaint, Grok Build, r/opencodeCLI, 2026-09-23: “if anthropic were truly zdr, this report wouldn't even be possible. the lab that shattered the zdr narrative was anthropic itself! you'd have a much better point saying that about grok build or zcode.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wo2tmb/deepseek_moonshot_kimi_xiaomi_under_investigation/pbjl3ju/)

- **One brand word, several separate subscriptions** ([Pricing and plan terms stated clearly and consistently](https://feedbackbench.com/criteria/billing.pricing_clarity.md), [IDE and editor integration](https://feedbackbench.com/criteria/setup.ide_integration.md)). The Grok name spans products billed by different companies, so users cannot tell which subscription unlocks which tool.
  Posts show users who pay for Grok, then hit a wall when a Grok-branded agent inside Cursor demands a Cursor subscription. One user lays out the split as a list and still asks whether they have misunderstood.
  
  Replies explain the split rather than fix it. Same brand word, different products, different bills. That confusion lands on the vendor's pricing pages, not on the user.
  Evidence:
  - Complaint, Grok Build, r/cursor, 2026-09-09: “the naming trips everyone. grok bot is a cursor product, so you need a cursor sub for it. grok imagine / regular grok chat is the separate grok subscription. same brand word, different products. if what you actually want is an ide with tools, cursor is the one that wires that up. grok build alone won't give you that surface.” [source](https://www.reddit.com/r/cursor/comments/1wbjv4t/im_confused_about_the_grok_and_cursor_plans/p8sm3pi/)
  - Complaint, Grok Build, r/cursor, 2026-09-09: “1. i use grok mainly for vdo work, with some vibe coding too. but when i want to use it like an ide, i don’t have tools such as codex, claude cowork, or antigravity, so i have to use grok build. 2. i also want to use grok bot, but it says i need a cursor subscription. i’m totally confused, especially because it’s called “grok” bot. so: * grok imagine → requires a grok subscription * grok bot → requires a cursor subscription is there another way? or am i misunderstanding the plans?” [source](https://www.reddit.com/r/cursor/comments/1wbjv4t/im_confused_about_the_grok_and_cursor_plans/)

- **Editor integration needs manual wiring** ([IDE and editor integration](https://feedbackbench.com/criteria/setup.ide_integration.md), [MCP servers, plugins, skills and hooks](https://feedbackbench.com/criteria/setup.extensions_mcp.md)). Grok Build ships without a native IDE surface, so users wire it in through custom MCP servers or simply switch to Cursor.
  All IDE-integration posts are complaints. A self-described fan wishes Xcode support were seamless instead of requiring a custom MCP server. Others tell newcomers to skip Grok Build and use Cursor, which already wires the Grok agent and image tools into an editor.
  
  The pattern: users like the agent but go elsewhere for the surface they work in.
  Evidence:
  - Complaint, Grok Build, r/vibecoding, 2026-09-07: “grok build user here. i prefer it, albeit i wish the xcode integration were seamless rather than having to do a custom mcp. just recently received claude code due to being a teacher and codex since i am a student back in university for four months.” [source](https://www.reddit.com/r/vibecoding/comments/1w74l2u/is_everyone_here_on_claude_or_codex_only_anyone/p8cx591/)
  - Complaint, Grok Build, r/cursor, 2026-09-09: “just use cursor instead of grok build. you can use imagine inside cursor by prompting the grok agent to call generateimage/imagine. also, codex and claude cowork aren't ides. codex/claude code are extensions you use inside an ide, generally either vscode or cursor. cowork is a different app.” [source](https://www.reddit.com/r/cursor/comments/1wbjv4t/im_confused_about_the_grok_and_cursor_plans/p8rkgrv/)

- **Patches fast, writes and plans poorly** ([Can do the user's kind of task](https://feedbackbench.com/criteria/work.capability.md), [Does unrequested work or over-engineers](https://feedbackbench.com/criteria/work.scope_overreach.md)). Users report Grok Build acts before thinking enough, favouring narrow local patches over architecture and producing prose they call unusable.
  One user quotes a review saying Grok 4.6 starts working before its reasoning is sufficient and reaches for local fixes on known problems. Another is leaving for Codex because Grok Build writes far worse than grok.com.
  
  A trial user puts it bluntly and would rather use GPT 5.4. Capability posts split close to evenly, so this is a live dispute, not a settled verdict.
  Evidence:
  - Complaint, Grok Build, r/codex, 2026-09-12: “i tried grok 4.6 via grok build cli for mac cuz they gave me 3 day free trial, is soooo bad, id rather use gpt 5.4 than grok.” [source](https://www.reddit.com/r/codex/comments/1we1a4j/tibo_tibo_tibo/p9aj19k/)
  - Complaint, Grok Build, r/codex, 2026-09-09: “i’m considering switching from grok build to codex. i need an ai that can write decently. i don’t need perfect writing or high quality writing. just natural, easy to read writing. grok.com is.. passable. but grok build is horribly horrible at writing. chatgpt.com is good, i’m hoping codex is passable. i don’t need perfection, i just need something that produces ok writing. either codex, claude code, or grok build (no online because i need it to read my files).” [source](https://www.reddit.com/r/codex/comments/1wb60ia/grok_to_codex_is_codexs_writing_decent/)
  - Complaint, Grok Build, r/ClaudeAI, 2026-09-21: “interesting experience. "grok 4.6 and grok cli respond very fast, but they often start working before the prior thinking is sufficient. it tends more toward making local patches on known problems, rather than actively improving the overall architecture" i have this exact problem with gemini as well. i trying to force is to think of general architecture over the local patches. but, it always reverses to easy and narrow patches. only claude models are able to do deep architectural analysis. it is interesting.” [source](https://www.reddit.com/r/ClaudeAI/comments/1vxzbij/after_using_claude_grok_46_and_gemini_37_flash_in/pb4qjfl/)

### What works

- **Generous limits for the price** ([How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md), [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md)). SuperGrok plans buy enough usage that users stop rationing, which several contrast with tiptoeing around Codex limits.
  Plan-value posts lean positive. One heavy-tier user says it is really hard to go over the limits. Another says Codex forces them onto smaller models to stretch quota, while SuperGrok gives near top-tier quality at roughly the same price.
  
  A $200-plan user says offloading worker jobs to Grok Build stretches their main plan to a full week.
  Evidence:
  - Praise, Grok Build, r/vibecoding, 2026-09-05: “grok build is great workhorse. with supergrok heavy it is really hard go over the limits. i was lucky enough to have cursor ultra which has great grok limits also and you can use anthropic models for harder/ui tasks. if i would not have the deal, then i would take cursor ultra or $100 subs from anthropic, codex and grok.” [source](https://www.reddit.com/r/vibecoding/comments/1w74l2u/is_everyone_here_on_claude_or_codex_only_anyone/p7x0u38/)
  - Praise, Grok Build, r/OpenAI, 2026-09-10: “while i love codex, you do have to be delicate with the limits, i see myself using luna a lot more than any other model just to preserve limits i got supergrok, and damn i'm just playing around with grok build, while not as feature rich as codex, for almost the same price, you get grok as your default model which, is sol level quality! with atleast terra level usage and for claude, you atleast get sonnet 5 which is better than luna, tagging it with /advisor is amazing so... sometimes i feel that with codex i'm comprising to make the limits feel better” [source](https://www.reddit.com/r/OpenAI/comments/1wc9wzb/switched_from_claude_to_codex_limits_feel_lesser/p8wbu5x/)
  - Praise, Grok Build, r/codex, 2026-09-23: “i am using 200 sub. it gets a lot better after i offload the worker jobs to another llm. i think there are many reasonable priced worker plans out there. should really try it. i think my plan can last for a week after getting grok build into my workflow. just wonder which is better, astra light or sol max.” [source](https://www.reddit.com/r/codex/comments/1wo746i/gpt6_astra_light_or_gpt6_sol_max_as_planner_and_qa/pbknkqk/)

- **Follows rules, stays in its lane** ([Direct in-prompt instructions and caps are followed](https://feedbackbench.com/criteria/context.instruction_following.md)). Users find Grok Build obeys instructions and avoids drift, doing exactly what is asked and nothing more.
  A new user calls it surprisingly better than Claude at following rules and not drifting into chaos. A Swift developer finds it solid for slightly-beyond-prototype work with a strong grasp of design principles.
  
  The same literalness has a cost. That user also says it is poor at anticipating issues, which echoes the shallow-patch complaints.
  Evidence:
  - Praise, Grok Build, r/ClaudeCode, 2026-09-22: “usage burns fast on codex too but at least is still competent on sol 5.6. claude on opus 5 has gotten nearly unusable. i will say that my first impressions of grok build are good. its surprisingly much better than claude at actually following rules and not drifting into pure insanity.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wmbt29/two_5x_sub_or_one_20x_sub/pba474l/)
  - Praise, Grok Build, r/google_antigravity, 2026-09-24: “i primarily work in swift so it's a mixed bag, especially during any transition. we're currently moving to ios 27, which some llms assert still doesn't even exist yet, so trying to do anything "new" is still best done by hand. i have started playing around with grok build for small personal projects i don't have time to work on but really want to tinker with and it's surprisingly good for slightly-beyond-prototype work. it has a strong grasp of design principles but it's very "dumb" when it comes to anticipating issues. it does exactly what you ask and nothing more.” [source](https://www.reddit.com/r/google_antigravity/comments/1wp6mpo/poll_how_do_you_code_in_late_2026/pbsymyj/)

- **Strong as a delegated subagent** ([Subagents, parallel agents and orchestrators](https://feedbackbench.com/criteria/work.multi_agent_orchestration.md), [Agent-performed code review finds real issues](https://feedbackbench.com/criteria/verify.agent_code_review.md)). Users slot Grok Build under a Claude or Codex orchestrator for research, worker jobs and adversarial review, where its speed and quota pay off.
  Multi-agent posts are all praise. Users name it their subagent and research engine while Claude handles clean work. One runs a delegate skill that pulls Grok 4.6 xhigh in as a reviewer and says it finds novel, high-value issues.
  
  Another runs it headless under a Fable orchestrator for cheap, cross-checked output.
  Evidence:
  - Praise, Grok Build, r/vibecoding, 2026-09-05: “codex or claude calling grok build at an agent works really well” [source](https://www.reddit.com/r/vibecoding/comments/1w74l2u/is_everyone_here_on_claude_or_codex_only_anyone/p7y2apg/)
  - Praise, Grok Build, r/vibecoding, 2026-09-04: “i use claude code for clean work, gpt for helping me prompting, deepseek for daily tasks, grok build for subagent & research workflow, gemini when polishing frontend ui/ux design since it's better than claude (at least for me).” [source](https://www.reddit.com/r/vibecoding/comments/1w74l2u/is_everyone_here_on_claude_or_codex_only_anyone/p7trtvo/)
  - Praise, Grok Build, r/ClaudeCode, 2026-09-22: “i’m sure there’s a more elegant way, i had astra leading, calling fable but it works fine in the reverse also. i created a ‘delegate’ skill and prompt the orchestrator agent to use the delegate skill to bring in whatever model(s) i specify. using claude -p when delegating to a claude model. i actually have an antigravity sub, a grok super heavy sub (which gives me quota via grok build and cursor ultra), and the new $50 muse sub. i always use claude or codex as the lead and then prompt situationally for them to bring in some combo of others via delegate. gemini flash 3.8 high, grok 4.6 xhigh, and muse 1.3 xhigh are all excellent adversarial reviewers and they all find novel high value things when used in that way.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wneic4/saw_this_today/pbgume6/)
  - Complaint, Grok Build, r/ClaudeCode, 2026-09-11: “i use fable as my orchestrator, sonnet recon and headless grok build. will shift to sonnet/opus build once the heavy grok discount runs out. it is slower, a lot slower but token efficient and all the models add a layer of checking the others work. less baby sitting and cleaner code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wd9eio/fable_as_orchestrator_and_opussonnet_as_executers/p9574gb/)

- **Fans call it top tier** ([Can do the user's kind of task](https://feedbackbench.com/criteria/work.capability.md), [Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md)). Enthusiasts report end-to-end task completion and hardware setup wins that beat their previous tools.
  One user says Grok 4.6 already beats Opus and Sol. A Flutter developer moved almost entirely to Grok Build because it finishes tasks end to end. Another credits it with configuring multi-GPU local hosting after months of being stuck, then pointing itself at those local models.
  
  Game builders like that it generates image assets as part of a build request.
  Evidence:
  - Praise, Grok Build, r/ClaudeCode, 2026-09-05: “grok 4.6 is already better the opus or sol. grok 4.7 comming in one week and being a fable class model together with grok build 100$ or 300$ plan going to be best of all subs. for me it already is” [source](https://www.reddit.com/r/ClaudeCode/comments/1w89cji/fable_vs_astra/p81o7ju/)
  - Praise, Grok Build, r/codex, 2026-09-05: “is anyone using astra/sol and grok 4.6 high via grok build? i haven’t received astra yet, but i do use sol. and as of late, i’ve almost entirely shifted to grok build (grok 4.6 high) - mostly because it’s way faster than gpt and very good at completing tasks end-to-end for my flutter project. when there is access to both llms, speed does get the veto from me, and i mostly use codex/gpt only as my general ai driver, and not for my flutter project anymore. how is it for others?” [source](https://www.reddit.com/r/codex/comments/1w7on0j/astra_is_absolutely_incredible/p7ys22g/)
  - Praise, Grok Build, r/LocalLLaMA, 2026-09-20: “i have 4 of the tesla v100 32gb cards running in my rig. something that i discovered is that the current version of grok build is uncannily good at setting up these cards tuning them selecting functioning models to download and getting it all up and running under lennox. i'm presently hosting three models qwen 3.8, qwen 3.6, and nemotron 3.5 with results that continue to surprise me. after i had grac set up the cards then i had grok build reconfigure itself to run using the cards it had just set up and it works just fine. it's not as fast'cause using rock 46 is the language model but it does get the job done your mileage might be might vary thought i would share this helped me get unblocked after a couple of months of screwing around with various problems on these cards and getting kind of frustrated.” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wl680s/finally_got_qwen_38_next_running_on_my_v100_6gpu/paw9a6m/)
  - Praise, Grok Build, r/vibecoding, 2026-09-04: “grok build does a great job, especially if you are looking to create assets for images or videos, like if you're building a video game. it generates characters with grok image as part of your request to build. really great. if you want to build in local models into those processes you can do that too.” [source](https://www.reddit.com/r/vibecoding/comments/1w74l2u/is_everyone_here_on_claude_or_codex_only_anyone/p7v9cyj/)

### Under the surface

- **Speed reports point both ways** ([Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md)). Some users pick Grok Build for raw speed while others call it very slow, so latency likely depends on model tier and routing.
  A Flutter developer says speed is why Grok Build won their workflow, and a review quoted by another user calls Grok 4.6 very fast. Against that, a Cursor user calls it cheap but very slow, and an orchestrator user lists slowness as the trade for token efficiency.
  
  Too few posts to explain the split.
  Evidence:
  - Praise, Grok Build, r/cursor, 2026-09-22: “i didn't find it very expensive when used in grok build, but it's very slow.” [source](https://www.reddit.com/r/cursor/comments/1wmswj9/grok_47_is_about_25_times_as_expensive_as_46/pbecymc/)
  - Praise, Grok Build, r/codex, 2026-09-05: “is anyone using astra/sol and grok 4.6 high via grok build? i haven’t received astra yet, but i do use sol. and as of late, i’ve almost entirely shifted to grok build (grok 4.6 high) - mostly because it’s way faster than gpt and very good at completing tasks end-to-end for my flutter project. when there is access to both llms, speed does get the veto from me, and i mostly use codex/gpt only as my general ai driver, and not for my flutter project anymore. how is it for others?” [source](https://www.reddit.com/r/codex/comments/1w7on0j/astra_is_absolutely_incredible/p7ys22g/)
  - Complaint, Grok Build, r/ClaudeCode, 2026-09-11: “i use fable as my orchestrator, sonnet recon and headless grok build. will shift to sonnet/opus build once the heavy grok discount runs out. it is slower, a lot slower but token efficient and all the models add a layer of checking the others work. less baby sitting and cleaner code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wd9eio/fable_as_orchestrator_and_opussonnet_as_executers/p9574gb/)

- **Discounts carry part of the adoption** ([How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md), [Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md)). Several users arrived through trials or heavy discounts and plan to reassess when promotional pricing ends, so plan-value goodwill may not survive repricing.
  One orchestrator user says they will move work back to Sonnet or Opus once the heavy Grok discount runs out. A trial user left unimpressed after three free days.
  
  Praise for limits is real, but some of it is tied to pricing users expect to change.
  Evidence:
  - Complaint, Grok Build, r/ClaudeCode, 2026-09-11: “i use fable as my orchestrator, sonnet recon and headless grok build. will shift to sonnet/opus build once the heavy grok discount runs out. it is slower, a lot slower but token efficient and all the models add a layer of checking the others work. less baby sitting and cleaner code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wd9eio/fable_as_orchestrator_and_opussonnet_as_executers/p9574gb/)
  - Complaint, Grok Build, r/codex, 2026-09-12: “i tried grok 4.6 via grok build cli for mac cuz they gave me 3 day free trial, is soooo bad, id rather use gpt 5.4 than grok.” [source](https://www.reddit.com/r/codex/comments/1we1a4j/tibo_tibo_tibo/p9aj19k/)

### Fine print

- All posts come from Reddit, and every area and criterion is flagged as too few posts.
- The cross-tenant session leak is one user's claim. The vendor reportedly closed it as a hallucination.
- Some praise comes from users on discounts or bundled plans, which may inflate plan-value sentiment.

## Top requests

What users ask to add or change, most asked first. 8 author-weeks ask for something. Requests do not change the Feedback Score. Rule: A separate pass by Claude Sonnet 5 reads every counted post and extracts what the author asks the agent or its vendor to add or change, with the criteria it maps to and a short normalised wording; it does not touch the labels or the Feedback Score. Claude Opus 5.5 groups the wordings within each criterion (the first criterion the request maps to) into themes; code counts them. A theme counts distinct author-weeks that ask for it, per agent; across agents, one author-week per agent. Themes asked in fewer than 2 author-weeks, and requests that share no theme, are not shown. Examples: up to 3 posts per theme from different authors, without slurs, preferring posts of 60 to 450 characters, most recent first.

No request is asked for in enough author-weeks to show.

## Facts

| Fact | Value |
|---|---|
| Version | v1.0 (out of beta 2026-08-07); underlying model grok-code-fast-1 |
| Released | Beta: 2026-05-14. v1.0: 2026-08-07 |
| Price | Bundled with SuperGrok Heavy ($300/mo); Grok Bot beta access also reachable via Cursor Ultra ($200/mo) or Cursor Teams Premium ($120/seat/mo) - standalone Grok Build pricing not clearly separated in sources |
| Model | grok-code-fast-1, trained from scratch (not the Grok 4 lineage), heavy on programming corpus and real-world PR post-training |
| Surface | CLI (local-first, no code sent to xAI servers) |

## Sources

| Channel | Source | Posts |
|---|---|---|
| Reddit | Posts that name it | 78 |

## Better than peers on

None.

## Worse than peers on

None.

## All 63 criteria

Criterion love: 0.5 is the category norm. n: rated author-weeks.

### Paying and limits: Too few posts (customer love 0.535, n 12)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [How much use a plan's price buys](https://feedbackbench.com/criteria/limits.plan_value.md) | Too few posts | 0.515 | 0.499–0.533 | 8 | 6 | 2 |
| [Single prompt, model or effort level consumes disproportionate quota](https://feedbackbench.com/criteria/limits.burn_rate.md) | Too few posts | 0.506 | 0.496–0.523 | 2 | 1 | 1 |
| [Pricing and plan terms stated clearly and consistently](https://feedbackbench.com/criteria/billing.pricing_clarity.md) | Too few posts | 0.499 | 0.496–0.500 | 1 | 0 | 1 |
| [Using an existing subscription across tools](https://feedbackbench.com/criteria/billing.subscription_portability.md) | Too few posts | 0.503 | 0.500–0.508 | 1 | 1 | 0 |
| [Short rolling usage window blocks or interrupts work](https://feedbackbench.com/criteria/limits.window_interrupts_work.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Price, allowance or plan terms changed](https://feedbackbench.com/criteria/limits.allowance_change.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Quota reset timing and bonus or banked resets](https://feedbackbench.com/criteria/limits.reset_schedule.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Usage meter visibility and accuracy](https://feedbackbench.com/criteria/limits.usage_meter.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Prompt cache hits, misses and invalidation](https://feedbackbench.com/criteria/limits.prompt_cache.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Pay-as-you-go overage, fallback billing and spend caps](https://feedbackbench.com/criteria/billing.overage_charges.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Free tier and free model availability and limits](https://feedbackbench.com/criteria/billing.free_tier.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-23, r/codex (Reddit): “i am using 200 sub. it gets a lot better after i offload the worker jobs to another llm. i think there are many reasonable priced worker plans out there. should really try it. i think my plan can last for a week after getting grok build into my workflow. just wonder which is better, astra light or sol max.” [source](https://www.reddit.com/r/codex/comments/1wo746i/gpt6_astra_light_or_gpt6_sol_max_as_planner_and_qa/pbknkqk/)
- Praise, 2026-09-22, r/cursor (Reddit): “i didn't find it very expensive when used in grok build, but it's very slow.” [source](https://www.reddit.com/r/cursor/comments/1wmswj9/grok_47_is_about_25_times_as_expensive_as_46/pbecymc/)
- Praise, 2026-09-11, r/ClaudeCode (Reddit): “i use fable as my orchestrator, sonnet recon and headless grok build. will shift to sonnet/opus build once the heavy grok discount runs out. it is slower, a lot slower but token efficient and all the models add a layer of checking the others work. less baby sitting and cleaner code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wd9eio/fable_as_orchestrator_and_opussonnet_as_executers/p9574gb/)
- Complaint, 2026-09-21, r/cursor (Reddit): “anyone want to share their ruleset for this pipeline? i am finding i get great results at about half the cost of a grok plan and build. here is what i am using. looking for ways to improve it and get it closer to a grok build quality. keep in mind, having robust project rules like core, ui, ui-chrome, schema, etc help a lot in first pass success. i find most with bad results don't have the right r…” [source](https://www.reddit.com/r/cursor/comments/1wmm80e/grok_47_xhigh_plan_composer_25_build_fast_mode_off/)
- Complaint, 2026-09-19, r/cursor (Reddit): “i've been using cursor for the past 10 months, first with the codex ide extension, and the last 45 days with everything else the same, but with the privoder switched to deepseek. the 8 billion tokens i've used over the past 45 days i paid us$**71.57** for - i spent about $800-900 for the previous 13-14 billion tokens with openai (lots of resets used judiciously). we're talking extremely cache heav…” [source](https://www.reddit.com/r/cursor/comments/1wk16qb/so_what_happened_to_cursor_in_the_past_few_weeks/papt0dm/)
- Complaint, 2026-09-15, r/cursor (Reddit): “hi there, i have a question regarding the grok bot usage limits: i currently pay \~30$ for my supergrok subscription, however i am a developer myself and consider getting cursor, i figured the 60$ cursor plan offers extended but not max grok bot usage. supergrok itself has some grok bot usage included, now i am curious if someone made a switch yet and noticed some usage improvement with grok bot?…” [source](https://www.reddit.com/r/cursor/comments/1whg7lf/should_i_stay_on_supergrok_or_switch_to_cursor/)

### Setting up and connecting: Too few posts (customer love 0.499, n 5)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [IDE and editor integration](https://feedbackbench.com/criteria/setup.ide_integration.md) | Too few posts | 0.493 | 0.485–0.500 | 3 | 0 | 3 |
| [Connecting own API keys, local models and custom endpoints](https://feedbackbench.com/criteria/setup.provider_byok_local.md) | Too few posts | 0.502 | 0.500–0.507 | 1 | 1 | 0 |
| [MCP servers, plugins, skills and hooks](https://feedbackbench.com/criteria/setup.extensions_mcp.md) | Too few posts | 0.502 | 0.500–0.507 | 1 | 1 | 0 |
| [Install, launch and sign-in](https://feedbackbench.com/criteria/setup.install_signin.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Onboarding, discoverability and documentation](https://feedbackbench.com/criteria/setup.onboarding_docs.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-24, r/ChatGPTCoding (Reddit): “all harnesses are tui. though a few like grok build has menus clickable by mouse. i actually recommend using that with a local model if you don’t want to subscribe to anything.” [source](https://www.reddit.com/r/ChatGPTCoding/comments/1womcvr/best_claude_code_alternatives/pboumxa/)
- Praise, 2026-09-10, r/cursor (Reddit): “try "grok build for vs code" extension. 120k+ installs via open vsx and vs code marketplace. open source. [<strict_link> works as a cursor ide extension too. unlike in cursor, you get oryginal harnesses for grok build, claude code, and codex. and bring your own subscriptions without cursor's. below, a remote control view, the ui is similar across all surfaces. <strict_link>” [source](https://www.reddit.com/r/cursor/comments/1wbust5/grok_officially_ruined_cursor_ide/p8wzb6g/)
- Complaint, 2026-09-09, r/cursor (Reddit): “just use cursor instead of grok build. you can use imagine inside cursor by prompting the grok agent to call generateimage/imagine. also, codex and claude cowork aren't ides. codex/claude code are extensions you use inside an ide, generally either vscode or cursor. cowork is a different app.” [source](https://www.reddit.com/r/cursor/comments/1wbjv4t/im_confused_about_the_grok_and_cursor_plans/p8rkgrv/)
- Complaint, 2026-09-09, r/cursor (Reddit): “the naming trips everyone. grok bot is a cursor product, so you need a cursor sub for it. grok imagine / regular grok chat is the separate grok subscription. same brand word, different products. if what you actually want is an ide with tools, cursor is the one that wires that up. grok build alone won't give you that surface.” [source](https://www.reddit.com/r/cursor/comments/1wbjv4t/im_confused_about_the_grok_and_cursor_plans/p8sm3pi/)
- Complaint, 2026-09-07, r/vibecoding (Reddit): “grok build user here. i prefer it, albeit i wish the xcode integration were seamless rather than having to do a custom mcp. just recently received claude code due to being a teacher and codex since i am a student back in university for four months.” [source](https://www.reddit.com/r/vibecoding/comments/1w74l2u/is_everyone_here_on_claude_or_codex_only_anyone/p8cx591/)

### Choosing models: Too few posts (customer love 0.510, n 2)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Which models are offered on a plan and when](https://feedbackbench.com/criteria/models.catalog_access.md) | Too few posts | 0.505 | 0.500–0.515 | 1 | 1 | 0 |
| [Quality got worse or better over time](https://feedbackbench.com/criteria/models.quality_drift.md) | Too few posts | 0.506 | 0.500–0.518 | 1 | 1 | 0 |
| [Automatic model routing and fallback](https://feedbackbench.com/criteria/models.routing_auto.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Reasoning effort setting and its defaults](https://feedbackbench.com/criteria/models.effort_control.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-24, r/opencodeCLI (Reddit): “<strict_link> from my experience, muse spark 1.3 at xhigh in opencode give me wrong answers all the time. it might flare better with muse code as the model is trained and refined around the harness, the same way grok inside opencode feels dumber compared to when it's inside grok build.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wopbtv/what_muse_spark_14_contributor_is_already_here/pbswa6q/)
- Praise, 2026-09-10, r/OpenAI (Reddit): “while i love codex, you do have to be delicate with the limits, i see myself using luna a lot more than any other model just to preserve limits i got supergrok, and damn i'm just playing around with grok build, while not as feature rich as codex, for almost the same price, you get grok as your default model which, is sol level quality! with atleast terra level usage and for claude, you atleast get…” [source](https://www.reddit.com/r/OpenAI/comments/1wc9wzb/switched_from_claude_to_codex_limits_feel_lesser/p8wbu5x/)

### Instructing and context: Too few posts (customer love 0.505, n 3)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Direct in-prompt instructions and caps are followed](https://feedbackbench.com/criteria/context.instruction_following.md) | Too few posts | 0.509 | 0.500–0.523 | 2 | 2 | 0 |
| [Memory and state carried across sessions](https://feedbackbench.com/criteria/context.session_memory.md) | Too few posts | 0.498 | 0.493–0.500 | 1 | 0 | 1 |
| [Persistent project rules files are read and obeyed](https://feedbackbench.com/criteria/context.instruction_files.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Asks the user versus guessing](https://feedbackbench.com/criteria/context.clarifying_questions.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Output degrades as the context window fills](https://feedbackbench.com/criteria/context.long_context_decay.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Context compaction keeps what matters, cheaply and quickly](https://feedbackbench.com/criteria/context.compaction.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Finding the right files in the codebase](https://feedbackbench.com/criteria/context.codebase_retrieval.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Images, PDFs and file attachments as input](https://feedbackbench.com/criteria/context.attachments.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-24, r/google_antigravity (Reddit): “i primarily work in swift so it's a mixed bag, especially during any transition. we're currently moving to ios 27, which some llms assert still doesn't even exist yet, so trying to do anything "new" is still best done by hand. i have started playing around with grok build for small personal projects i don't have time to work on but really want to tinker with and it's surprisingly good for slightly…” [source](https://www.reddit.com/r/google_antigravity/comments/1wp6mpo/poll_how_do_you_code_in_late_2026/pbsymyj/)
- Praise, 2026-09-22, r/ClaudeCode (Reddit): “usage burns fast on codex too but at least is still competent on sol 5.6. claude on opus 5 has gotten nearly unusable. i will say that my first impressions of grok build are good. its surprisingly much better than claude at actually following rules and not drifting into pure insanity.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wmbt29/two_5x_sub_or_one_20x_sub/pba474l/)
- Complaint, 2026-09-08, r/codex (Reddit): “the way i have learned to see it after 3500 hours of experience with vibe coding is that its best to treat all models, whether it's codex, claude code, grok build etc, like a dumb employee that can work hard and comes up with something good every now and then, but you need to manage this employee a lot and if you don't steer it, it will start creating a lot of overhead, over-engineer things that a…” [source](https://www.reddit.com/r/codex/comments/1wamtly/i_dont_find_building_with_codex_or_any_ai_easy_at/p8leyzb/)

### Doing the work: Too few posts (customer love 0.513, n 14)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Can do the user's kind of task](https://feedbackbench.com/criteria/work.capability.md) | Too few posts | 0.501 | 0.486–0.516 | 10 | 6 | 4 |
| [Subagents, parallel agents and orchestrators](https://feedbackbench.com/criteria/work.multi_agent_orchestration.md) | Too few posts | 0.504 | 0.500–0.510 | 2 | 2 | 0 |
| [Does unrequested work or over-engineers](https://feedbackbench.com/criteria/work.scope_overreach.md) | Too few posts | 0.499 | 0.496–0.500 | 1 | 0 | 1 |
| [Spins, loops or gets stuck without progress](https://feedbackbench.com/criteria/work.stuck_loops.md) | Too few posts | 0.499 | 0.496–0.500 | 1 | 0 | 1 |
| [Safety filters block legitimate coding tasks](https://feedbackbench.com/criteria/work.safety_refusals.md) | Too few posts | 0.509 | 0.500–0.527 | 1 | 1 | 0 |
| [Frontend and visual UI output](https://feedbackbench.com/criteria/work.frontend_ui.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Diagnosing and fixing reported bugs](https://feedbackbench.com/criteria/work.bug_diagnosis.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Breaks existing code or reintroduces bugs](https://feedbackbench.com/criteria/work.regressions_introduced.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Stops mid-task or answers instead of acting](https://feedbackbench.com/criteria/work.premature_stop.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Long unattended runs and goal/loop mode](https://feedbackbench.com/criteria/work.long_running_autonomy.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Games checks instead of fixing the problem](https://feedbackbench.com/criteria/work.reward_hacking.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Risky or irreversible actions without confirmation](https://feedbackbench.com/criteria/work.destructive_actions.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Git commits, branches and sync](https://feedbackbench.com/criteria/work.git_workflow.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Computer use and browser control](https://feedbackbench.com/criteria/work.computer_browser_use.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Tool approval prompts and autonomy modes](https://feedbackbench.com/criteria/work.permission_prompts.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Plan-before-edit mode](https://feedbackbench.com/criteria/work.plan_mode.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Length and clarity of replies, summaries and comments](https://feedbackbench.com/criteria/work.response_verbosity.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Caves to or argues with the user's judgement](https://feedbackbench.com/criteria/work.sycophancy_pushback.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-20, r/LocalLLaMA (Reddit): “i have 4 of the tesla v100 32gb cards running in my rig. something that i discovered is that the current version of grok build is uncannily good at setting up these cards tuning them selecting functioning models to download and getting it all up and running under lennox. i'm presently hosting three models qwen 3.8, qwen 3.6, and nemotron 3.5 with results that continue to surprise me. after i had g…” [source](https://www.reddit.com/r/LocalLLaMA/comments/1wl680s/finally_got_qwen_38_next_running_on_my_v100_6gpu/paw9a6m/)
- Praise, 2026-09-13, r/cursor (Reddit): “its wierd seeing the many bad experiences from all of you with grok 4.6. i actually had t opposite experience. i was unhappy with the high prices of claude enterprise in my company. bei g the guy responsible for rolling out ai to everybody i was looking for alternatives. i started to try grok build and cursor and after initially having a problem with trusting their models in any way i became prett…” [source](https://www.reddit.com/r/cursor/comments/1weddl3/thats_has_happened_to_cursor/p9i8zid/)
- Praise, 2026-09-07, r/opencodeCLI (Reddit): “i don't use kimi k3. but i can tell you that glm 5.3 flash > muse spark 1.3 xhigh > qwen 3.8 flash. qwen 3.8 flash makes mistakes with confidence and is also slow, takes a lot of detours, and makes you spend twice as many tokens despite being "cheaper." muse spark 1.3 xhigh is intelligent but very lazy; it's a terrible agent to work with. it forgets to call tools and always looks for the quickest…” [source](https://www.reddit.com/r/opencodeCLI/comments/1w5zhbn/kimi_k3_vs_glm_53_vs_qwen_38_max_vs_muse_spark_13/p89e49c/)
- Complaint, 2026-09-21, r/ClaudeAI (Reddit): “interesting experience. "grok 4.6 and grok cli respond very fast, but they often start working before the prior thinking is sufficient. it tends more toward making local patches on known problems, rather than actively improving the overall architecture" i have this exact problem with gemini as well. i trying to force is to think of general architecture over the local patches. but, it always revers…” [source](https://www.reddit.com/r/ClaudeAI/comments/1vxzbij/after_using_claude_grok_46_and_gemini_37_flash_in/pb4qjfl/)
- Complaint, 2026-09-18, r/codex (Reddit): “objectively, no. anthropic's only good model is fable. you can only use 50% of your usage on it. their other models are both bad and extremely overpriced. sonnet costs like 9-15x more per task than luna. luna max actually performs similar to opus on medium. so you get like 20x more work done with luna than with opus but 50% of your subscription is basically locked to opus. and it's $100, not $20.…” [source](https://www.reddit.com/r/codex/comments/1wjhw4l/is_claude_a_value_switch_now/paj18pd/)
- Complaint, 2026-09-12, r/codex (Reddit): “i tried grok 4.6 via grok build cli for mac cuz they gave me 3 day free trial, is soooo bad, id rather use gpt 5.4 than grok.” [source](https://www.reddit.com/r/codex/comments/1we1a4j/tibo_tibo_tibo/p9aj19k/)

### Checking and finishing: Too few posts (customer love 0.506, n 2)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Agent-performed code review finds real issues](https://feedbackbench.com/criteria/verify.agent_code_review.md) | Too few posts | 0.503 | 0.500–0.509 | 2 | 2 | 0 |
| [Claims work is done or fixed when it is not](https://feedbackbench.com/criteria/verify.false_completion.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Builds, tests or runs its own changes](https://feedbackbench.com/criteria/verify.self_testing.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Reviewing and approving the agent's changes](https://feedbackbench.com/criteria/verify.change_review_ui.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-22, r/ClaudeCode (Reddit): “i’m sure there’s a more elegant way, i had astra leading, calling fable but it works fine in the reverse also. i created a ‘delegate’ skill and prompt the orchestrator agent to use the delegate skill to bring in whatever model(s) i specify. using claude -p when delegating to a claude model. i actually have an antigravity sub, a grok super heavy sub (which gives me quota via grok build and cursor u…” [source](https://www.reddit.com/r/ClaudeCode/comments/1wneic4/saw_this_today/pbgume6/)
- Praise, 2026-09-11, r/ClaudeCode (Reddit): “i use fable as my orchestrator, sonnet recon and headless grok build. will shift to sonnet/opus build once the heavy grok discount runs out. it is slower, a lot slower but token efficient and all the models add a layer of checking the others work. less baby sitting and cleaner code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wd9eio/fable_as_orchestrator_and_opussonnet_as_executers/p9574gb/)

### Interface and sessions: Too few posts (customer love 0.503, n 1)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [How the interface shows work, and what the user can configure](https://feedbackbench.com/criteria/ui.display_settings.md) | Too few posts | 0.504 | 0.500–0.512 | 1 | 1 | 0 |
| [Saving, switching, resuming and rewinding sessions](https://feedbackbench.com/criteria/ui.session_history.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Stopping and steering a running agent](https://feedbackbench.com/criteria/ui.interrupt_steer.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Mobile, remote-control and voice access](https://feedbackbench.com/criteria/surfaces.remote_mobile.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Cloud and remote sandbox execution](https://feedbackbench.com/criteria/surfaces.cloud_sessions.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-20, r/codex (Reddit): “i really want to use codex cli but i find it hard to follow because it doesn't hide all the thinking and edited files like it does on the desktop app. claude code has focus mode and grok build hides that stuff by default. is there a way to do that in codex cli?” [source](https://www.reddit.com/r/codex/comments/1wlffmo/is_codex_itself_really_the_best_harness_to_use/paynsfa/)
- Praise, 2026-09-20, r/codex (Reddit): “i really want to use codex cli but i find it hard to follow because it doesn't hide all the thinking and edited files like it does on the desktop app. claude code has focus mode and grok build hides that stuff by default. is there a way to do that in codex cli?” [source](https://www.reddit.com/r/codex/comments/1wlffmo/is_codex_itself_really_the_best_harness_to_use/paynu2a/)
- Praise, 2026-09-20, r/codex (Reddit): “focus mode on codex cli? i really want to use codex cli but i find it hard to follow because it doesn't hide all the thinking and edited files like it does on the desktop app. claude code has focus mode and grok build hides that stuff by default. is there a way to do that in codex cli?” [source](https://www.reddit.com/r/codex/comments/1wli2pv/focus_mode_on_codex_cli/)

### Reliability and speed: Too few posts (customer love 0.503, n 3)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Latency, throughput and fast mode](https://feedbackbench.com/criteria/rel.response_speed.md) | Too few posts | 0.500 | 0.492–0.509 | 3 | 1 | 2 |
| [Outages, server errors and capacity or rate errors](https://feedbackbench.com/criteria/rel.service_errors.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Client crashes, freezes and failed tool execution](https://feedbackbench.com/criteria/rel.client_failures.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Updates break working setups](https://feedbackbench.com/criteria/rel.update_breakage.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Praise, 2026-09-05, r/codex (Reddit): “is anyone using astra/sol and grok 4.6 high via grok build? i haven’t received astra yet, but i do use sol. and as of late, i’ve almost entirely shifted to grok build (grok 4.6 high) - mostly because it’s way faster than gpt and very good at completing tasks end-to-end for my flutter project. when there is access to both llms, speed does get the veto from me, and i mostly use codex/gpt only as my…” [source](https://www.reddit.com/r/codex/comments/1w7on0j/astra_is_absolutely_incredible/p7ys22g/)
- Complaint, 2026-09-22, r/cursor (Reddit): “i didn't find it very expensive when used in grok build, but it's very slow.” [source](https://www.reddit.com/r/cursor/comments/1wmswj9/grok_47_is_about_25_times_as_expensive_as_46/pbecymc/)
- Complaint, 2026-09-11, r/ClaudeCode (Reddit): “i use fable as my orchestrator, sonnet recon and headless grok build. will shift to sonnet/opus build once the heavy grok discount runs out. it is slower, a lot slower but token efficient and all the models add a layer of checking the others work. less baby sitting and cleaner code.” [source](https://www.reddit.com/r/ClaudeCode/comments/1wd9eio/fable_as_orchestrator_and_opussonnet_as_executers/p9574gb/)

### Account and support: Too few posts (customer love 0.494, n 4)

| Criterion | Reading | Customer love | 95% interval | n | Praise | Complaint |
|---|---|---|---|---|---|---|
| [Data retention, training use and deployment isolation](https://feedbackbench.com/criteria/account.data_privacy.md) | Too few posts | 0.496 | 0.491–0.500 | 3 | 0 | 3 |
| [Support, refunds and issue handling](https://feedbackbench.com/criteria/account.support.md) | Too few posts | 0.499 | 0.496–0.500 | 1 | 0 | 1 |
| [Wrong charges, failed payments and plan provisioning](https://feedbackbench.com/criteria/account.billing_errors.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |
| [Account bans and access restrictions](https://feedbackbench.com/criteria/account.bans_restrictions.md) | Too few posts | 0.500 | 0.500–0.500 | 0 | 0 | 0 |

Most recent posts:

- Complaint, 2026-09-23, r/opencodeCLI (Reddit): “if anthropic were truly zdr, this report wouldn't even be possible. the lab that shattered the zdr narrative was anthropic itself! you'd have a much better point saying that about grok build or zcode.” [source](https://www.reddit.com/r/opencodeCLI/comments/1wo2tmb/deepseek_moonshot_kimi_xiaomi_under_investigation/pbjl3ju/)
- Complaint, 2026-09-22, r/AI_Agents (Reddit): “yeah and guess what, xai closed this as model hallucination. the bug bounty report was in for about a month or so, now they're not using this model anymore for grok build..” [source](https://www.reddit.com/r/AI_Agents/comments/1wexro6/fyi_malicious_actors_could_likely_hijack_your/pba84bx/)
- Complaint, 2026-09-18, r/codex (Reddit): “i’ve done a bit of research. i think if you use grok through cursor, not grok cli, it uses cursor’s mechanisms of getting the agent to do the task. not grok’s. and cursor’s privacy is much much higher. if your argument is more about morality than your codebase’s privacy, that’s also completely valid.” [source](https://www.reddit.com/r/codex/comments/1wjhw4l/is_claude_a_value_switch_now/pamwwss/)
