@merniti
iAccount based inUnited States
About this account
- Account based in
- United States
- Connected via
- United States App Store
Account-level information from X, not a live location or the device used for a specific post.
founder at @beam_cloud (YC W22). I like hackathons, container runtimes, and steak medium rare
NYC / SF
Joined June 2017
- Tweets1.3K
- Following1K
- Followers4.2K
- Likes11.4K
Pinned Tweet
7am flight out of SFO
- Leave house 1h before flight leaves
- 13 min Uber to airport, views of the Bay
- TSA agent smiles asks if you like the Grateful Dead
- Ads for AI agents that cure cancer
- Multiple food options spanning global cuisines
- 2 min walk to gate
- Get upgraded to business class
- Depart 5 min early
7am flight out of JFK
- Leave house 3h before flight leaves
- 73 minute Uber to airport, bumper to bumper traffic
- TSA agent hates you
- Ads for underwear
- No food options besides Jamba Juice and hardboiled eggs
- 17 min walk to gate
- Get downgraded to seat next to bathroom
- Depart 2 hours late
was curious how open weight alternatives compare to Jev, so rewrote this demo with SemIf instead (basically open source Jev)
Jev is already insanely cheap, somehow this is still 50% cheaper
Jev (@typesafeai) is so insane & cheap for search!!
> 6000+ @ycombinator Startups indexed.
> Sub 1 second search results.
> 90M tokens & $2.7 in total testing costs.
Search any startup in a second, in any way!
- Color - Niche - Your Competitor - Age - Image - etc...
> watch the entire video, it's so freaking cool omg!
> this is the coolest thing i have ever built for fun! (worked on it for 2 days straight!)
try all the open source jev alternatives
beam.cloud/playground
life after removing the free trial from your AI SaaS app, blocking IPs from Indonesia, and requiring every signup to add a credit card
jev was released less than a week ago, and there are already tons of OSS alternatives
SemIf is basically as good as jev, costs 50% less, and has no waitlist
try all the open jev alternatives
beam.cloud/playground
honestly crazy how meta just shipped openclaw and added $200B in market cap overnight
everyone underestimated just how much compute the labs would consume
each company at the AI cloud layer (in the middle of the sandwich) earns >90% of their revenue from labs
there’s vanishingly little capacity left for anyone else
WHO GETS PAID IN THE AI CLOUD STACK
AI clouds are becoming the middle layer between scarce AI infrastructure and companies that actually need compute:
Who supplies the AI clouds
• $NVDA sits at center of compute layer by supplying the GPUs that determine how much capacity these platforms can actually bring online.
• $MU & $SKHY supply HBM required to keep those GPUs fed as model sizes and inference workloads continue scaling
• $DLR, $EQIX, $CORZ & $APLD provide physical data center infrastructure underneath the cloud layer giving AI clouds another path to scale without owning every building themselves.
Who turns that infrastructure into compute
• $CRWV, $NBIS & $IREN sit directly in middle of the stack by combining GPUs, power, networking and software into usable AI compute that customers can rent.
Who buys the compute
• $MSFT, $META & $GOOGL are unique because they are both customers and competitors by renting external capacity when internal supply is constrained while continuing to build their own
• $SHOP & $CRWD show where next leg can come from as enterprise demand broadens customer base beyond a handful of hyperscalers and AI labs.
• OpenAI & Anthropic represent AI lab customer base where demand can scale really quickly as training and inference requirements grow.
jev is very exciting because a ton of LLM tasks don’t actually need a paragraph of text generated
we’re throwing LLMs at all sorts of problems, many of which are just classification. and we’re spending millions of dollars on RL to make LLMs write better answers when their actual job is making a decision
but jev doesn’t generate tokens, it just makes decisions
the entire AI economy is priced around token generation, and we just realized that a ton of inference spend is being used to generate tokens we never even needed
past two weeks i’ve seen multiple teams go from bleeding money on opus to printing money just by moving to deepseek, or going from deepseek to fine-tuned qwen
it's starting to feel like a trend