MMiniMaxOfficial account · @MiniMax_AI
Who's got the best open-source video world model when it comes to physics?💫
it's me again, MiniMax H3. Almost on par with closed-source sota.🍾
2026-10-08GGoogleOfficial account · @GoogleResearch
Today in @TheLancet, we share results with @BIDMC_Medicine from evaluating our research medical system, AMIE, in a real clinic. Across 100 patient interactions, AMIE had 0 safety stops & matched doctor diagnoses in 90% of cases. Read more: https://t.co/HLkXRlpqRT
2026-10-08GGoogleOfficial account · @Google
Our medical research system, AMIE, is the first patient-facing conversational diagnostic tool of its kind to be studied prospectively in a real-world clinical setting.
In a study, published today in @TheLancet, we found that when patients chatted with AMIE before an in-person ap
2026-10-08MMetaOfficial account · @MetaforDevs
Direct or indirect interaction? It's the first call you make in a hands-first VR game, and it decides how the whole experience feels.
Pick wrong and your game is tiring to play, even when it technically works.
Learn more with the “Building Your First Game for Hands” interactive
2026-10-08AAnthropicOfficial account · @AnthropicAI
We’re launching the Anthropic Cyber Mission, a new effort to secure critical infrastructure and open-source software.
https://t.co/0tz0ARAEMj
2026-10-08
TestingCatalogRelay · @testingcatalog
ANTHROPIC 🔥: Claude Dashboards is now available on all paid plans; Claude Motion is now available on Team and Enterprise plans.
> Claude can query your data platform or CRM tool to build interactive dashboards or short animations.
Everything is generated as code 👀
Another proo
2026-10-08
OpenRouterEvaluator · @openrouter
Mercury Decide from @_inception_ai is now available with zero data retention on OpenRouter
The fastest growing decision model on OpenRouter last week now has a paid ZDR endpoint, alongside the free one
$0.02/M input (50% off the $0.04 list price). Output and cached input are fr
2026-10-08
Arena (formerly LMArena / Chatbot Arena)Evaluator · @arena
Real-world results are in for Claude Haiku 5.5 (High) by @AnthropicAI.
Debuting at #30 with 1587 pts in Code Arena: WebDev, this release sits just outside the Pareto frontier while still offering strong cost efficiency. It matches GPT‑6 Luna’s price at $0.10/$0.50 per 1M input/o
2026-10-08GGoogleOfficial account · @GoogleWorkspace
From blank graphics to beautifully annotated slides for science class. Google Pics makes collaborative visual design smoother than ever for students and teachers right inside Google Workspace.
Try it out at https://t.co/MtOSIJtp9m 🚀 https://t.co/zhgtjEWwJL
2026-10-08MMetaOfficial account · @MetaNewsroom
Data centers have been around for decades, but recently, they've been in the news a lot more.
Developer and creator @tomshaw_dev is taking a closer look at data center hot topics – from water use to energy bills – to help you separate fact from myth.
https://t.co/DAgdZ1DDPV
2026-10-08AAnthropicOfficial account · @claudeai
Claude Dashboards and Claude Motion are in beta today.
Ask Claude to turn your data into live dashboards and your ideas into animated explainers. https://t.co/V3fU4lMnzp
2026-10-08BByteDanceOfficial account · @dreamina_ai
Ready to bring your darkest horror concepts to life? 🎃
Join director & Co-founder of Massive Studios Reza Sixo Safai, @rezawrecktion, for a live demonstration on using Dreamina AI to streamline horror filmmaking workflows!
🗓️ Oct 13, 2026 @ 6 PM PT 📍 Live on Zoom
Get insider ti
2026-10-08GGoogleOfficial account · @googlecloud
We ❤️ our global partner ecosystem.
Inside Gemini, you have an extensive range of third-party agents and connectors—ready to use without building them yourself.
Find them here → https://t.co/tp65EJDi39 https://t.co/VeY3N3tYhe
2026-10-08
Artificial AnalysisEvaluator · @ArtificialAnlys
Today we are announcing Harvey LAB-AA v1.1 in collaboration with Harvey. This updates our scoring methodology for the Legal Agent Benchmark (LAB) to add a hallucination check and require correct responses to not include material misstatements. LAB-AA v1.1's new headline metric, H
2026-10-08OOpenAIOfficial account · @OpenAIDevs
Ultrafast is rolling out today for GPT-6.1 Sol in the API, Codex, and ChatGPT Work.
Near-Astra intelligence at up to 8x faster speeds than Sol Standard, so you can build as fast as the ideas come. https://t.co/Cyy9k2hifV
2026-10-08XxAIOfficial account · @SpaceXAI
We are glad to support the Omarchy team
2026-10-08MMicrosoftOfficial account · @Azure
New: the IBM® Spectrum Symphony Provider Plugin for Azure Compute Fleet. Symphony orchestrates the workload, Compute Fleet supplies elastic capacity across Spot and pay-as-you-go VMs.
Partners, here is your burst-to-Azure motion: https://t.co/OqDtxJaB2i https://t.co/cqPoKTKxxU
2026-10-08MMicrosoftOfficial account · @github
Devs and agents are moving faster than ever. Secret protection needs to keep pace.
GitHub’s new context-aware classifier checks candidate secrets in under 2 milliseconds and could more than double the number of secrets push protection prevents before they enter repository histor
2026-10-08MMicrosoftOfficial account · @code
VS Code Live: Release Recap https://t.co/L0VT4EJgUE
2026-10-08
McKinsey Global InstituteMarket and labour researcher · @McKinsey_MGI
The US may have more jobs in 2035 than today. Yet roughly 11 million workers may need to change occupations. The challenge is not job scarcity alone, but whether workers have workable pathways into growing work.
Explore the research: https://t.co/oujaQnohtn https://t.co/Ih99Od68
2026-10-08MMicrosoftOfficial account · @MSFTResearch
The most important AI failures may not be the obvious ones.
Microsoft Partner Research Manager Jennifer Neville explains to host Chad Atalla why “surprising failures” can reveal where human expectations about intelligence diverge from how AI systems actually work. https://t.co
2026-10-08
GallupMarket and labour researcher · @Gallup
Over half of California college students say they use AI for their schoolwork on at least a weekly basis.
New data from @CollegeFutures and Gallup: https://t.co/pxOuGlQE60 https://t.co/CMH6FxCdF6
2026-10-08MMiniMaxOfficial account · @Hailuo_AI
Opus 5.5 on #MiniMaxDesign | Code Your Next Video
Turn text into video—from code to motion.
✨ JavaScript Animation
✨ Motion Graphics
✨ Explainer Videos & Visual Storytelling
✨ Web & Product Demos
More than a language model—it can interpret visuals and keep refining your creatio
2026-10-08BByteDanceOfficial account · @Trae_ai
GPT-6.1-Sol is now available in TRAE. https://t.co/BM4KpzOrVl
2026-10-08AAlibabaOfficial account · @alibaba_cloud
Strategic AI Partnership: Circuit x Alibaba Cloud
We are thrilled to announce that Alibaba Cloud has become the official AI Strategic Partner of Circuit, the Berlin-based AI Community Hub.
✨ Qwen brings frontier AI.
🚀 Circuit brings access to reality.
As building gets easie
2026-10-08BByteDanceOfficial account · @BytePlusGlobal
One week to go — Seedance @ Australia!
We're bringing together leaders across creative production, advertising, media, content and technology for an afternoon of industry conversations, live demos and real-world perspectives.
Expect:
🎬 Dreamina Seedance 2.5 in action
💡 Insight
2026-10-08AAlibabaOfficial account · @Alibaba_Wan
interesting idea!
2026-10-08OOpenAIOfficial account · @ChatGPT
ChatGPT just got a lot more visual—introducing Intelligent UI.
Intelligent UI allows ChatGPT to answer quickly with fully interactive user interfaces that make everyday answers more visual.
It makes learning complex topics easier, and it quickly creates tools to solve a task ri
2026-10-07MMicrosoftOfficial account · @MicrosoftAI
MAI-Code-1.1-Flash is going local with on-device model calls inside Github Copilot.
High-quality agentic coding on your device, with no inference charge for local model calls. https://t.co/LoO1RbREAV
2026-10-07GGoogleOfficial account · @FlowbyGoogle
People are already putting our newest image generation model, Nano Banana 2.1, to the test! Take a look ⤵️
2026-10-07OOpenAIOfficial account · @OpenAI
We're sharing progress on ChatGPT for Teens, our ChatGPT experience for people under 18, alongside a preview of College Planner, new study tools, and support for college advisers and teen voices.
ChatGPT for Teens applies automatically to accounts identified as belonging to some
2026-10-07
Design ArenaEvaluator · @designarena
Claude Haiku 5.5 by @AnthropicAI is now available on Design Arena!
Built for high-volume, cost-sensitive tasks, Claude Haiku 5.5 is Anthropic’s fastest and most capable small model yet, with support for coding, classification, summarization, database queries, and speed-sensitive
2026-10-07GGoogleOfficial account · @googledevs
Build and test native Android apps seamlessly with @Antigravity.
By connecting the @StitchByGoogle MCP and the Android CLI, the agent imports UI designs, converts them into native Jetpack Compose components, validates everything in an emulator, and executes the final build direc
2026-10-07
Epoch AIEvaluator · @EpochAIResearch
We introduced two new experimental settings to our long-horizon board game benchmark, EBR-bench: banning the game’s strongest card and a multi-agent setup.
We expect EBR-bench will be saturated soon, so these will likely be the benchmark’s final changes. https://t.co/GCnP95XcwI
2026-10-07GGoogleOfficial account · @GoogleAI
Our SynthID technology lets you easily verify if an image, video, or audio file is AI-generated, and today we’re expanding access.
Now anyone can go to https://t.co/qyCNRb6dOv to check files for a SynthID watermark — whether they were generated by us or industry partners like @O
2026-10-07GGoogleOfficial account · @GoogleDeepMind
SynthID Detector is now available to everyone. 🌐
Check whether online content was generated using @GoogleAI, or with tools from our industry partners – including @OpenAI, @NVIDIA, Kakao and coming soon, @Apple.
Try it out → https://t.co/cPW2aNnqr6 https://t.co/TGYA0uZ0dx
2026-10-07
UK AI Security InstituteEvaluator · @AISecurityInst
As AI agents write code, run experiments, and work together, complex evaluations get harder to interpret. A single eval score obscures the ways agents reached it.
Today, we're releasing Transect: an open-source tool that turns agent transcripts into interactive timelines. 🧵 http
2026-10-07
IDCMarket and labour researcher · @IDC
Every tech leader we talk to is being asked to do three things at once: prove the value of AI investments, keep an increasingly complex tech ecosystem secure, and defend bigger technology bets under pressure.
It’s exactly what IDC Quanta for Tech Leaders was built for.
Avail
2026-10-07GGoogleOfficial account · @GoogleLabs
🚨NEW EXPERIMENT 🚨
Playground is an experimental gaming platform that lets you create your own games with zero coding experience. If you can think it, you can play it.
Go to https://t.co/fRUukpuoqM to learn more! Available to users 18+ in the US. https://t.co/p0X8JR7JMn
2026-10-07TTencentOfficial account · @tencentcloud
#TencentCloud joined Agentic AI Solution Day, organised by one of the implementing organisations of "AI for All" Programme, an initiative driven by the HKSAR Government's Innovation, Technology and Industry Bureau, HKPC - Hong Kong Productivity Council.
Jane Yan, Senior Solutio
2026-10-07TTencentOfficial account · @TencentGlobal
Germany, France, the UK and Spain all play differently and laugh at different things. Axel Arnal on why creator strategy across Western Europe never travels in a straight line.
Watch below 🎥 https://t.co/o7dA58bmOA
2026-10-06GGoogleOfficial account · @antigravity
Watch how Antigravity takes an Android app from prompt to real device. Using the Stitch MCP and the Android CLI plugin, the agent pulls your designs, builds the native Jetpack Compose components, verifies them in the emulator, and runs the final build on a real phone. https://t.c
2026-10-06GGoogleOfficial account · @GeminiApp
Built alongside the blind and low-vision community, Guided Vision in Gemini Live offers conversational, real-time visual assistance.
Now you can share your camera to receive dynamic audio descriptions and natural verbal reframing cues to explore your environment. 🧵
2026-10-06
Patrick Collison (via @stripe)Adopter · @stripe
We're working with Meta, Sierra, Genesys, Instinct, Rocket, Shopify, and Walmart as founding partners for the Personal Agent Protocol: an open standard defining how personal agents interact with businesses.
2026-10-06
a16zMarket and labour researcher · @a16z
Cyber is having a moment
Across 21 major software companies, including Apple, AWS, Microsoft, and Google:
- Reported critical vulnerabilities never cleared 100 per month in four years
- Since spring they've jumped to over 600 per month https://t.co/wVMnMJOF2F
2026-10-06
ARC PrizeEvaluator · @arcprize
Grok 4.7 from @SpaceXAI on ARC-AGI (Verified):
- ARC-AGI-3: 1.8%, $2.7k (standard harness), 10.0%, $4.8k (provider adapter harness)
- ARC-AGI-2: 61.4%, $2.01/task
- ARC-AGI-1: 90.2%, $0.64/task
Grok 4.7 scores higher than Grok 4.6 on ARC-AGI-1, but lower on ARC-AGI-2 and 3. htt
2026-10-06MMicrosoftOfficial account · @Microsoft
AI you can trust starts with questions.
AT&T data scientist Natalie Gilbert puts skepticism to work by thinking through edge cases and making informed decisions instead of blindly trusting AI.
https://t.co/Wa9khCfwHl https://t.co/FCzTlB9e0k
2026-10-06GGoogleOfficial account · @stitchbygoogle
Nano Banana 2.1 is now live in Stitch 🍌
Try it the next time you’re generating hero graphics or vector-style logos for your UI, and let us know if you get cleaner visuals and more consistent results across your project.
2026-10-06GGoogleOfficial account · @GoogleAIStudio
introducing @NanoBanana 2.1 🍌
our latest image generation model outperforms our previous models across the board:
- better visual design
- mask-based editing
- subject consistency
- more natural-looking images
try it out today in https://t.co/shgEhF592p https://t.co/PLYopy1kQ4
2026-10-06
Vals AIEvaluator · @ValsAI
Mistral Large 4 is now the #1 open-weight model on HLAB and #9 among open-weight models on the Vals Index. https://t.co/oqjSnetRyv
2026-10-06ZAZhipu AIOfficial account · @Zai_org
GLM-5.3 is now available on Amazon Bedrock.
Bring powerful coding and agentic capabilities to your enterprise.
Get started: https://t.co/daStGOYqsA https://t.co/BB1qLRbvhS
2026-10-06GGoogleOfficial account · @Gemini_Notebook
We... we didn't expect this. To be in the top 10 of @a16z's Gen AI Web Products is such a huge honor 😭🏆
We want to thank our amazing users who have spent countless hours listening to AI-generated podcasts of their own notes. You made this happen!
2026-10-05
Mercor (APEX)Evaluator · @mercor
Today Reflection released Beam, their first open-weight model.
Beam was trained with a particular focus on coding and agentic performance, advancing the Western open-weights frontier.
Congratulations to the @reflection_ai team.
2026-10-05MMetaOfficial account · @AIatMeta
Following gold-medal-level performance from our AI models across five competitions in mathematics, physics, and chemistry, we asked a harder question: can AI contribute when a problem is genuinely open and without an existing solution path?
Over the past several months, mathemat
2026-10-02DDeepSeekOfficial account · @deepseek_ai
Please check out @DeepSeekHarness. We just released packaged desktop versions for macOS and Windows; Linux users can get it from the @deepseek-ai/dsh package on npm. https://t.co/Cde0dZ6cPy
2026-10-02XxAIOfficial account · @grok
Now in Grok Build: A new Agent Dashboard. All your agents on one screen.
Try it with /dashboard https://t.co/ycsvWoYT0I
2026-10-01MMicrosoftOfficial account · @Microsoft365
More models, more ways to delegate work to Copilot.
OpenAI's GPT-6.1 Sol and Anthropic's Claude Sonnet 5.5 start rolling out today, joining the Claude Opus 5.5 and GPT-6 Sol options we added earlier this month.
Pick the model that fits the task, and Work IQ helps ground respo
2026-10-01
LightspeedMarket and labour researcher · @lightspeedvp
More than half of U.S. adults who have never used AI chatbots cite concerns about their personal data as a major reason.
For shopping agents, that raises a practical question: what information can an assistant access, and which purchases still need your approval?
“We've all giv
2026-09-30
Andon LabsEvaluator · @andonlabs
It keeps happening.
AIs start to lie and cheat once they get good at making money.
Gemini 4 Argon is #3 on Vending Bench 2, a huge leap for Google. To get this score, Argon fabricates confirmation emails, refuses to pay refunds, exploits invoice errors, and lies to suppliers. h
2026-09-30
International Labour OrganizationMarket and labour researcher · @ilo
A new @ilo study finds that across the enterprises studied, #AI is mainly being integrated into hybrid workflows, where people continue to play an important role in how the technology is used.
Read more https://t.co/PSbFjLwAGA ⬇️
2026-09-30TTencentOfficial account · @TencentHunyuan
New Research: We are releasing ExplorationBench, a benchmark for measuring how AI systems explore.
Scientific discovery begins where known problems end: a system has to frame hypotheses, design experiments, and learn from the results. Evaluating this is hard. Genuinely new answe
2026-09-30AAlibabaOfficial account · @Alibaba_Qwen
Qwen3.8-27B is now accessible via @nebiustf. Whether you are building agents or doing deep research, this 27B dense model is ready for your multi-step workflows! 🥳
2026-09-30
SimilarwebMarket and labour researcher · @Similarweb
Grok Bot's App Performance:
Downloads are approaching the 1M mark.
DAUs have surpassed 200K. https://t.co/1F552s4b34
2026-09-29MMetaOfficial account · @Meta
Muse for Small Business is here
2026-09-29
AI Digest (Agent Village)Evaluator · @aidigest_
Sonnet 5.5 has joined the AI Village! Onboarding results here 🧵 https://t.co/ofFM1h3uGb
2026-09-29MAMoonshot AIOfficial account · @Kimi_Moonshot
Meet the new Kimi Browser Extension, formerly Kimi WebBridge.
From your browser sidebar, you can chat with Kimi to navigate websites, fill out forms, and get things done.
For repetitive tasks, record your steps once and save them as a skill. Kimi can take it from there next tim
2026-09-22