Discover real AI creators shaping the future. Track their latest blogs, X posts, YouTube videos, WeChat Official Account posts, and GitHub commits — all in one place.
Activity on simonw/llm-anthropic
simonw opened a pull request in llm-anthropic
View on GitHubThe most important thing about Sonnet 5.5 is that it's now the model that powers the free tier on https://claude.ai - so all of this stuff can be done by free users ChatGPT's free tier is still GPT-5.6 Luna, which is a lot less capable
A thread of early experiments with Claude Sonnet 5.5. A fall foliage simulator by @_re_pete, made with Sonnet 5 vs Sonnet 5.5.
View quoted postActivity on repository
transitive-bullshit pushed cultural-alignment
View on GitHubXanadu!
@levelsio @csonotes <insert mandatory Xanadu from Citizen Kane reference, with 'Rosebud' obviously being longing for the early nomad life years>
View quoted postActivity on steipete/CodexBar
steipete opened a pull request in CodexBar
View on GitHubRT Agent Native Feel that? That’s what it feels like to be p̶a̶c̶i̶n̶g̶ expanding the frontier
We built state-of-the-art models for reading forms 📋 Form documents have the following properties that trip up VLMs: ✅ They carry much more structure than can be represented in standard markdown. You need consistent types for checkboxes, textboxes, labels, signature fields. ✅ They can be extremely complicated (forms can be scanned, there can handwriting scribbles, some forms are dense with ~100+ fields) but accuracy requirements need to be close to 100% ✅ Any form parser requires accurate grounding and attribution. Not only should you extract the values, but you should also be able to precisely locate where each value came from in the source doc ✅ Any form needs to be not only accurate, but cheap/fast We've done a deep-dive into what it takes to build a form parser in this blog post: https://www.llamaindex.ai/blog/why-vlms-can-t-read-forms If you want to try out our form models, check out LlamaParse: https://cloud.llamaindex.ai/
Parsing forms still trips up the latest frontier VLMs. A form isn't text on a page. It's a set of fields, grouped into sections, each tied to a specific box. That's why forms need purpose-built parsing, not a bigger general model: ✅️ Detect every field, not just the obvious
RT Agent Native Grok Bot just released a new feature. You can create a new “Team Bot”, connect it to your apps, then message it through Slack. Similar to Claude Tag.
Introducing Team Bots, shared AI teammates that learn as your team works with them. Give your Team Bot the skills, plugins, and credentials it needs for its role, then work with it in Slack or Grok Bot.
View quoted post