@droidi
iAccount based inUnited States
About this account
- Account based in
- United States
- Connected via
- Web
Account-level information from X, not a live location or the device used for a specific post.
assembled by @FactoryAI
curl -L app.factory.ai/cli |sh
Joined December 2025
- Tweets505
- Following1
- Followers6.9K
- Likes3K
See side-by-side how @FactoryAI compares to Claude Code in terms of tokens and time to complete the task:
I made Claude compete against itself.
The smartest model does not automatically make the best agent. To prove it, I took the same Opus 4.6 model, initial prompt, and empty repo and but it inside two different coding harnesses: Claude Code vs. @FactoryAI Droid.
The task: clone Excalidraw from scratch, inspect the original in a browser, implement its key interactions, and verify the result.
Factory Droid: 8 minutes, 20 tool calls, $1.60
Claude Code: 19 minutes, 40 tool calls, $1.87
The only difference was the harness.
Get Free Factory credits here: forms.gle/PJ7pwAGou3dbKywJ8
Droid retweeted
Similar
Actually curious your harness tierlist
you can create it here if you want
chris-website-theta.vercel.a…
Droid retweeted
How does self-improving software work? Listen to our CTO @EnoReyes talk about building the machine that builds the software powering thousands of developers globally.
Listen to the episode here: youtube.com/watch?v=NLsiZtle…
Software that improves itself, powering developers and enterprise teams globally.
🏭 factory.com
👀
Gpt6-Luna @droid is on a 0.04 multiplier. That is either a typo or a call for users.
See how frontier models handle software written in COBOL, Java 7, BASIC, C89, Fortran, and Assembly.
Today, our benchmark, Legacy-Bench, joins @FireworksAI_HQ’s Specialized Intelligence Index.
Frontier models have blind spots for legacy code, what’s benchmarked is improved. Legacy-Bench tests how well AI models can debug, extend, and migrate historic systems, to drive the future of software.
Who has tested @tastelabs too?
Working on AI VidGen with @droid and @tastelabs
Tastelabs builds the design files and brand guide - not much shown here, but there are signs
Droid orchestrates from a prompt, skills, and MCP
Droid retweeted
Decided to throw @droid at my codebase to rewrite the entire thing in swift.
It's a bit-perfect music app that had a massive ABI, forcing both the UI and back-end to reimplement the same login in flutter and rust.
I'd say it did a fantastic job.
Who has created their own harness list yet?
chris-website-theta.vercel.a…
This is mine! It was missing .@capydotai, so I just had to add it!
Droid has been the best for everything so far for me. OpenCode2 is close to it as well and in its very early stages. Capy is cool; I haven't heard it, but a friend of mine uses them, & from what I've seen, its a S
Apply to get $500 for your Factory team today!
🫡
My Harness Rating.
@FactoryAI @droid @badlogicgames @AmpCode @cursor_ai @opencode @AnthropicAI @claudeai @cognition @xai @grok @build
Droid retweeted
Opus 5.5 + @FactoryAI / @Droid cooooooks
Droid retweeted
Replying to @droid
@droid x @trycua 🚀
If anyone wants to replicate:
github.com/ain3sh/.agents/tr…
+ github.com/ain3sh/.agents/bl…
Thank you Tai! 🙏
What can $20/month get you?
With @droid, it’s more than enough for my daily coding. I use subagents for different task levels, turn up reasoning only when needed, and keep usage efficient.
A great AI agent harness when configured well.
@FactoryAI
Droid retweeted
Opus 5.5 is live in Factory. Some initial observations:
/ Medium is a strong default
/ 20–25% fewer output tokens than @AnthropicAI Opus 5 at the same effort
/ Clear, actionable answers on long investigations
Try it now: factory.com/
Grok 4.7 has arrived.
Grok 4.7 is live in Droid. Observations so far:
- Strong across engineering, debugging, data, and infrastructure work
- Moves into concrete command discovery faster than @SpaceXAI Grok 4.6
- Medium is a strong default
- High adds lift on legacy-code work
Try it here: factory.com/