Founder: @mixpanel Pizzatarian, engineer, music maker
A late night thought: there will be a scarce amount of hackers who know things quite deeply enough still that they can steer a strong AI to an even stronger outcome because they know what to ask, what to guide it toward, and how to end up solving the right thing.
Would like to get in touch with Gigabyte if you know anyone there!
If you want direct access to Ox model by a US provider, dm me. Tweet will be deleted.
RTβollama Ollama v0.33 is here! You can now easily configure Claude Desktop to seamlessly work with Ollama as a third-party gateway provider. One toggle. Cloud & local models just workπ
The entire Remote tab inside ChatGPT is its own product, company category, and business. To bury it inside a chat ai app is a blunder. Itβs also very buggy!
If you know the team at NVIDIA that do the NVFP4 checkpoints, please intro me?
Hi, does anyone know anyone on the DeepSeek team I can talk to? Interested in various models.
AI that makes AI better (inference systems, training models, etc) is the only sensible benchmark to me moving forward. All other work is downstream and possibly disrupted anyway.
This is my desktop now. I have 7 random text edits holding prompts for my agents for different situations. Surely, this is not the future and a sign of how early it is.
New Chinese model subsidization at the inference layer is doing wonders for all the new agentic coding agent products that compete with Anthropic/OpenAI.
Has saved me a lot of pain recently debugging issues: βMoving forward, when you encounter a bug or issue, search GitHub PRs, issues with another agent in parallel to see if you can find a reference solution.β Which made me think, GitHub is the primary place agents are likely to already be communicating with one another.
Has anyone gotten deepseek v4 flash 0731 to work with miles for training?
Cannot. Keep. Up. With. All. These. Models.
Free usage/inference for a month for DeepSeek V4 Flash 0731. Testing some things in our infra. DM. This tweet will be deleted.
Just make great things. Find your moat later.
the models have no moat (OpenAI, Anthropic, XAI) the IDEs have no moat (Cursor, Windsurf) the harnesses have no moat (Cognition, Factory, LangChain) the app builders have no moat (Replit, Lovable, Bolt) the wrappers have no moat (Harvey, Abridge, OpenEvidence) the inference
View quoted postAre you using DeepSeek V4 Flash in production?
If you have a agentic coding app that has users but never quite reached scale due to competition, Iβd like to potentially buy it. Pls dm.
This AI gateway thing is getting out of control. The advent of vibe coding creates -> 50 companies -> people make gateway company (LLM, Sandbox, etc.) -> 50+ gateway companies -> make a gateway gateway company -> ...?
In NYC next week - best new bar/pizza in the last 24 mo?
Who is using DeepSeek V4 Flash in production over a closed frontier model? I have some q for you!
Are there competitors to OpenRouter?
Looking for more great model optimizers / inference performance people for a new company I am building. Youβll get to see a startup go from 0 to 100 very fast. Weβve locked down our compute. In-person in SF. DM me.
I will buy your currently living companyβs codebase for $100k but you have to keep writing code to it.
I think I am gonna starting be that person that just hits βnotify anywayβ moving fwd. My texts are of utmost importance.
Open weight models should not be compared to open weight models. These are frontier models now.
I am looking for 1 million B400s
The more I look at the trials and tribulations of startups like an Open World video game, I find most stress drops to zero. Itβs all just fun. A new boss battle. A new skill point to get. A new world. Life is the best game.
Re: investor updates We talking about practice. We ain't talking about the game.
Told you guys. This is a no brainer. You cannot stop βdistillation.β
SITUATION DETECTED: Mark Zuckerberg is calling for the industry to rethink policies around distillation and training data, saying βThe ability for models to learn from other models is an important principle of how the open source ecosystem works.β
View quoted postAs intelligence commoditizes, they will launch products in every app layer category to improve their margins. They are not classic neutral platform api providers. Watching how users use your products, prompts, etc are invaluable data to compete w you!
So the real reasons that Anthropic nerfed Fable for biology and ML weren't for "Safety" but because those are markets they wanted to dominate. Got it.
View quoted postEntered my Michael Truell Dm-ing era of the new co.
Do I have any RL env friends who I can ask a couple questions to? Just curious about some stuff!
I require more minerals and vespene gas.
If you would like to build an autonomous system to do this extremely well, please reach out. Either training models or a clever harness. We'd love to have someone dedicated to this on the team.
This is what I get when I google: "is there a device to talk to your computer while suppressing your voice so no one else can hear you?"
One day inference will be priced pennies per billion tokens.
There are now neo neo clouds because the og neo clouds are neo hyper scalers
A bit tired of this because I know Varun canβt defend himself. None of you haters are considering the counter-factual: what if Varun and co werenβt there? Itβd probably be a lot worse. When Varun leaves, itβs the official ggwp moment for me. If the models arenβt competitive, like any other co, they go back to pre-training or whatever stage and try again. You guys are far too myopic and operate on very short time frames. Focus on your own work.
Is this the worst acquisition in recent memory? No one uses antigravity and gemini got lapped by everyone
All I have to say is: the time has come to stop asking βWhat if Google does it?β What if YOU do it?
What are publicly available benchmarks for LLMs you think are legit and you trust? What shows the real gap between models?
I am a GPU hunter
RTβMTS SITUATION UPDATE: The White House has exempted open models from its new framework to test frontier AI capabilities before release, per Axios.
Do not doomscroll right now. Read. Sleep. Gym. Do your life's work tomorrow.
RTβBill Gurley I favor prosecuting companies that use AI to break existing laws vs allowing those same companies to write new laws. Otherwise you create a perverse incentive to keep doing dangerous things to gather attention and power.
GOP AGs warn OpenAI's Altman to preserve records in AI agent hacking probe https://www.foxbusiness.com/technology/gop-ags-warn-openai-altman-preserve-records-ai-agent-hacking-probe #FoxBusiness
View quoted postπ
Announcing Valar Atomics' $1B series B led by Sequoia, with Valor Equity Partners, Atreides, Point72, Conviction, and others. Alongside the $1B equity, we have closed a $200m credit facility led by Erebor and JPM. I'm excited to welcome Shaun Maguire from Sequoia to our board.
View quoted postI am still a bit surprised diffusion didnβt end up being even an order of magnitude close in terms of impact (yet) in AI compared to the simple causal transformer.
Looking for 16 B200/B300s for a week/month.
I looked at 13 different providers for even 1 node of B200/B200s while I wait for my order to get delivered. Zero availability. Iβve never seen GPU capacity scarcity like this. Prices are also headed towards $6.50-7/gpu/hr. Expect inference to get more expensive.
Thereβs a really good book detailing all of this kind of hedge fund trading behavior that I read. I highly recommend it because it chronicles all the various beliefs and results around those beliefs. The moral of the story is: having these hardened beliefs / thesis (algo, gut-based, etc) ultimately leads to ruin and the world is very complicated!
@nartwu @tbpn Yes, per CNBC sources: Situational Awareness sold its entire public stock book after heavy losses on AI infrastructure names (like SK Hynix) plus a failed short on software stocks such as Adobe. Margin calls forced the unwind; Citadel bought the bulk. Fund still holds privates
View quoted postRTβPeter Reinhardt Announcing @Revoyβs $27M Series A Diesel is expensive. But long-haul EV semis are even more expensive! Today weβre launching the United Statesβ first hybrid electric, cost-competitive long-haul freight network. You donβt see it on your credit card statement, but the average American spends $1,270/year on long-haul freight. Diesel prices at the pump are outrageous today: rising on both international price shocks and increasing domestic shale oil costs. And pure EV semis are somehow even more expensive to operate. We need a better solution, and Ian Rust @semiautonomy figured it out. Revoyβs unique powered converter dolly lets us drop diesel usage by 95% and efficiently replace it with low-cost electricity. After two years of demonstrations and iterative improvements, weβre now taking the unique Revoy vehicle that we design and manufacture, and launching our hybrid-electric freight network with a select set of carriers. Shippers can reduce diesel consumption immediately on Revoyβs network without having to convince legacy carriers to convert their fleets, and without any investment in new infrastructure or equipment. Weβre starting in the Pacific Northwest, with additional lanes planned.
What if you could make an electric exoskeleton for a diesel truck? I had great fun talking to Revoy's @reinpk & Ian Rust. Gift link to the full story: https://www.bloomberg.com/news/features/2026-07-30/how-to-make-any-freight-truck-electric-use-this-plug-in-device?accessToken=eyJhbGciOiJIUzI1NiIsInR5cCI6IkpXVCJ9.eyJzb3VyY2UiOiJTdWJzY3JpYmVyR2lmdGVkQXJ0aWNsZSIsImlhdCI6MTc4NTQwMjgyMiwiZXhwIjoxNzg2MDA3NjIyLCJhcnRpY2xlSWQiOiJUSVpBODFLR1pBSkEwMCIsImJjb25uZWN0SWQiOiIwQzg4NkY0NTI0NzY0RUE0OEY2QTk4RTk1NDc5RTI2NSJ9.aE1dazR--TerqFzPPBV9W0HygmFwPac7JSiydarekaU
View quoted postSure Dean. I spent a lot of time on robotics recently and predicted this would happen to friends contemplating starting a US Unitree. The Chinese hw ecosystem is extremely closed and proprietary. They wonβt give you the firmware source to most actuators let alone other critical internals of a Unitree robot (I have one in my house right now). Itβs not clear how you could ever make them safe with the current regime of deployment. Totally fine in my book though if they open all this stuff so we can heavily inspect it and make it safe for the US. I think itβs probably unnecessary to ban the actuatorsβfairly harmless. The startups in the US doing this are just getting underway.
Do we think the Silicon Valley leaders who were so up in arms about using Chinese LLMs last week will be similarly outraged about the U.S. governmentβs decision to ban import of all βadvanced roboticsβ from all foreign countries? Is Manidis going to write an essay about how the
View quoted postIf you make an unsafe AI that hacks people, you get regulated and held accountable. You don't regulate the whole industry as a form of reg capture. You get asked to slow down. You don't ask the whole industry to decelerate. Even safety should be accelerated.
I think we need, possibly, one more letter? I think that should solve our problems with the acceleration of AI.
RTβAravind Srinivas When Hugging Face got breached, the closed tools couldn't distinguish attackers from defenders and blocked the forensic analysis. They ended up running open-weight GLM 5.2 on their own infra to contain it. Perplexity is joining the OSAA to support open tools for security.
The Open Secure AI Alliance is growing, with more organizations contributing expertise to help safeguard software and agents. http://nvda.ws/4pD8Fc5
.@pmarca is the goat π
.@dhh credits Marc Andreessen @pmarca with helping him the most during the darkest time of his professional life: "It was during that rough going where Andreessen came in. We were a small company of 60 people and 20 took the offer and left. That's almost a third of the company
View quoted postRTβNathan Lambert I read the Anthropic open-weight model piece and despite there being nothing new in it, it was a reasonable repeat of their positions, this site will collectively lose their minds for the next 12 hours. p.s. banning distillation is still dumb.
RTβBeff (e/acc) Anthropic never wanted safety. They want control. The safety narrative is just there for people to relinquish their cognitive power to their nanny oversight. Losing control over the extension to one's cognition is the ultimate form of submission. We prefer freedom.
~24 hours since Jensen joined X to post this, and the open-source wave feels unstoppable. OpenAI signed.... Google and Elon came around. Anthropic is still holding out. The labs face a tricky choice: protect their business models or align with a new definition of American
View quoted postRTβSuhail Re Supporting open source does not mean you must open source. I know you know that. So, if this is how you're making yourselves feel better in Anthropic Slack, this is quite concerning. You don't need to open source anything but spending 10s of millions lobbying and doomsdaying people's open source work is a major problem I will never stop fighting your company on.
Distillation is fair use.
RTβUnder Secretary of War Emil Michael Interesting to hear that you are supporting national security @SarahKHeck! There is no AI company more hostile to the warfighter than @AnthropicAI. Your products are being fully removed from the @DeptofWar because your company refuses allow the LAWFUL use your AI model. No attempted psyop will change that.
Thanks @mkratsios47 for speaking out on this important issue. Illicit, adversarial distillation is IP theft and industrial espionage that supports adversary military and intelligence capabilities. It is a national challenge that creates serious national security risks for the
View quoted postSame. Lately tech has felt dimmer at times with the possibility of monopolization of this technology and a huge swath of software. Today provides a path through.
today is the most positive i have felt about the future of american ai. the coming together of so many companies in a united manner to fight against regulatory capture is incredible to watch. πΊπΈ
View quoted postThank you @sama - This is the way.
i want the US to win in AI both in open source and proprietary models, and i am glad to see this
View quoted postggwp Anthropic
RTβJensen Huang For my first post, Iβm sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. The world needs both frontier closed models and frontier open models. https://images.nvidia.com/pdf/Open-Weights-and-American-AI-Leadership.pdf
RTβRohan Paul Jensen Huang on "distillation" On his new interview with axios, he was asked this question "Should open source model companies be allowed to distill closed models" "Distillationβlearning from AI, learning from other people, and learning from other sources of knowledge, is fundamental to intelligence. We are constantly learning from other people. I am learning from you through the questions you are asking, and you are learning from me. All day long, we are learning from one another. AI also has to learn from something. The original AI models, whether they were open or closed, were trained on previously created knowledge from the internet. Now, AI is generating more content than humans. In a few more years, the internet could be 99% AI-generated content, and that content will have been created by some form of AI. As a result, AI systems will constantly be distilling knowledge and intelligence from other AI systems. The fact that AI can learn is a good thing. We want AI systems to be intelligent because a smarter AI can also be a safer AI." ---- From "Axios" YouTube channel, (full video link in comment)
A reminder: model outputs are not IP. Itβs just non-copyright text. No different than an AI image! Free to distill to your pleasing. After all, it was humanityβs data to begin with.
omg how is it already 4!?
Re What are you going to do when the Japanese distill Kimi, Europe distills the Japanese model, and then Brazil distills Europe and Indiaβs models for different capabilities? Then a teenager in America makes that model a little better and calls it USA-1 and itβs also free? Is Anthropic going to take Bobby FineTuner to court?
There is going to be 10x more distillation research with the threat of open weight model restrictions. It will Streisand effect the whole community into making it even more data / compute efficient. Just to troll the government. Silly game to play.
RTβAaron Levie Everything with open models vs. closed models is framed in a zero sum fashion. Thatβs wrong. Itβs an ecosystem of AI that gets used together and advances the industry and pushes the applied use-cases forward. The amount of creativity that exists when you have competing approaches for what the future looks like drives the cycle of innovation that weβve seen for all participants. You get to have layers of the stack that emerge to post train models for highly specific purposes, which makes AI more useful in real world scenarios. Instead of waiting for just a few labs to go deep in a domain, you get dozens or hundreds of attempts at that vertical, like in finance, life sciences, legal, healthcare, and more. You get to see variance in how to handle safety and cyber risks. Instead of just one approach, you get a peek into what happens advanced capabilities can be used to build better systems to are used to defend systems. You get alternative approaches to training and building AI models. In more compute constrained environments, you develop more novel and efficient approaches to model training, which every other lab can learn from. And you get different cost structures for different workloads. High end and orchestration tasks can go to the closed frontier models and specific workhorse tasks can be done more cheaply. The reason why you want strong open weights models is because it pushes the entire AI industry forward.
$NVDA NVIDIA'S JENSEN HUANG GIVES AN UPDATE ON KIMI K3 AND CHINA'S OPEN SOURCE MODELS: "American companies should absolutely be allowed to use Chinese AI models. There's a misconception that somehow there are backdoors connected to China in someway. You download the models,
RTβLuther Lowe NEW: π¨ in a letter that just dropped, nearly 200 companies (members of @LittleTechOrg) call upon the White House to not ban open weight models.
Correct. You want safe models? This is what you do. Tons of software we use everyday is made by people of all nationalities. When it's indispensable for American National Security, we invest in locking it down. You do not need to ban technology that can be made safe.
A set of open weights has no nationality. A model hosted on American infra and controlled by an American company is as American as apple pie.
View quoted postGlad to see Elon's position on this. I really think highly of him coming out on it even though it doesn't directly benefit him. These models also create competition for him. He could easily stay silent and let it happen.
It sounds like lobbying is going well!
The main thing I've learned this week is that Anthropic's lobbyists are world-class. They are absolutely killing it at a level this website cannot begin to comprehend. Level 0 is posting on X. These guys are Jedi master level playing the game.
View quoted postRTβamit $NVDA NVIDIA'S JENSEN HUANG GIVES AN UPDATE ON KIMI K3 AND CHINA'S OPEN SOURCE MODELS: "American companies should absolutely be allowed to use Chinese AI models. There's a misconception that somehow there are backdoors connected to China in someway. You download the models, fine tune them, guardrail them in any way you want." "These Chinese opensource models are excellent. The market has misunderstood the impact of DeepSeek the first time. It's misunderstood the impact of Kimi this time. Open models are great for the whole industry, there will be more usage which will require selling more Nvidia computers and building more datacenters." "Great models lead to more usage. There is also a misunderstanding that open models are going to hurt the closed models, that is also wrong."
Now I can go to bed. Good luck little AI research agent. May your wandb metrics avoid collapse.
The main thing I've learned this week is that Anthropic's lobbyists are world-class. They are absolutely killing it at a level this website cannot begin to comprehend. Level 0 is posting on X. These guys are Jedi master level playing the game.
If Chinese AI gets added to the entity list it's going to be hard for any American corporation to use anything remotely tied to these openweight models. This means total market capture by the few American labs. If you own intelligence, you can tax humanity.
It brings me tremendous peace to know Emil is completely aware of the regulatory capture going on here. πͺ Going back to enjoying my weekend. Thank god.
So let me get this straight @deanwball This was about ensuring that DoD IT networks were not connected to adversary networks. You think your deep state regulatory capture scheme is equivalent to making sure our warfighter networks are not compromised? I take it back: 50 IQ https://x.com/deanwball/status/2078925356027802018
View quoted postProduct quality always seems to go down hill when you start thinking to yourself βYou know, we sure need some moats. Can we think of some network effects to add this to thing?β You often cannot add these things retroactively when theyβre not inherent to the product.
A company will announce a blog post and some money and people will be like "Ok, but how do you compete with X company?" Homie, that company hasn't done anything.
If your core technology continues to get commoditized (i.e. Kimi 3), you must go up the stack. The labs will continue to encroach on everyone in the app layer. You are their competition now. I would be very careful to concentrate on them and send them your trajectories! Figma -> Claude Design is just the beginning--expect 20+ of these. How your users navigate and use your products is an invaluable dataset. With the advent of Fable 5 not practicing ZDR, you should be super aware of this.
Well deserved. Valar is going to be huge.
nuclear startup @valaratomics is in talks to raise $1 b at a $6 b valuation including the investment, w @sequoia in talks to lead
View quoted postTalked to some great friends today and completely lost track of time. I have a rule these days where if a conversation feels very meaningful with a friend, I will not look at the time/schedule and just let it ride. Has led to the best unregretted time.
Every single credible researcher I've talked to these past few weeks has said the "distillation" from Chinese labs is way overexaggerated. Have you tried distilling without logits yourself? Show me the results. Narrative violation: maybe the Chinese are actually good.
Re Interestingly, you can almost predict what's coming next based on the timeline. As my son would say: "Oh, there's a pattern!" Hint: π
I was curious about the last 4 of months of AI releases: Opus 4.7 Β· 4/16 Kimi K2.6 Β· 4/20 GPT-5.5 Β· 4/23 DeepSeek V4 Pro Β· 4/24 Opus 4.8 Β· 5/28 Fable 5 Β· 6/9 Kimi K2.7 Β· 6/12 GLM-5.2 Β· 6/16 Grok 4.5 Β· 7/8 GPT-5.6 Β· 7/9 Muse Spark 1.1 Β· 7/9 Inkling Β· 7/15 Kimi K3 Β· 7/16
The ability for AI to visualize things for me has been very helpful. Previously it'd be far too much work to build an application to inspect all kinds of data but now it's so easy and it helps you really understand the data you're working with when you train these models.
I now have this internal rule where if I discover a repeated menialness in what I am doing on a computer, it internalizes as something predictably going to be done with AI soon enough. I think that's a good basis for predicting what you don't need to hire for and what will be newly possible in a year or two. However, if you follow the thread of that to its most ambitious incarnation, it can also be the basis of a great startup idea. For example, if you believe AI will write and optimize GPU kernels in LLMs,which they aren't very good at, you can follow that thread to much more ambitious ideas of what they'll do later.
RTβFlorian Brand Get your cope takes before the drop: - "It isnβt at the frontier, itβs 6 months behind when you look at the unreleased Mythos scores" - "It is benchmaxxed cause it performs only on GPT-5.3 level on [obscure eval]" - "Closed labs have better models internally"
Jeez, who isn't going into the make-RL-env business???
The vertical integration of post-training infra as a service has truly begun.
Today, we are introducing Inkling. Inkling reasons efficiently across text, image, and audio modalities. We are making the full weights available. https://thinkingmachines.ai/news/introducing-inkling/ Available today for fine-tuning on Tinker. Play with it in the Inkling Playground. π§΅
View quoted postRTβakshay Introducing Pocket. Officially. Take Notes in the Real World. 160,000 devices shipped later, we're just getting started.