XX / Twitter166
◈Bluesky60
YHacker News58
r/Reddit11
@Mastodon6
project rollup
Default
301
mentions
across 18 keywords
across 18 keywords
net sentiment
+25%
107 positive
161 neutral
33 negative
volume & sentiment over time
May 13 – Sep 28, 2026 · 3-day buckets
sources
most discussed
claude / pricing
active days
53
mentions
Showing 101–107 of 107 mentions
-
Google's statement that it includes a watermark that can identify them as AI-generated is completely braindead as to how disinformation works. You think people spending literally three seconds to derail a discussion give a shit about that? This is so easy to abuse www.404media.co/google-earth... Google's statement that it includes a watermark that can identify them as AI-generated is completely braindead as to how disinformation works. You think people spending literally three seconds to derail a discussion give a shit about that? This is so easy to abuse www.404media.co/google-earth...view source ↗
-
The Kimi K3 Moment> tried Kimi K3 on a task I've done with every other model I use regularly and found it chewed a lot longer on the problem and ate up almost the entirety of a 5 hour usage limit on their $19 plan ArtificialAnalysis puts Kimi K3 just below DeepSeek v4 & GLM 5.2 in token use per task, which is about 2x to 3x more tokens … > tried Kimi K3 on a task I've done with every other model I use regularly and found it chewed a lot longer on the problem and ate up almost the entirety of a 5 hour usage limit on their $19 plan ArtificialAnalysis puts Kimi K3 just below DeepSeek v4 & GLM 5.2 in token use per task, which is about 2x to 3x more tokens than Grok 4.5: https://x.com/ArtificialAnlys/status/2077832879187620192 / https://archive.vn/zBbFi 2 other open weights MiMo v2.5 & MiniMax M3 are comparatively thrifty. > Subscription usage limits are hard to measure as none of the providers tell you directly what it means in terms of tokens or anything else you can easily compare I always put my coding subscriptions (that allow it) through "AI gateways" (Cloudflare & OpenRouter are free) which help track token use. In my experience, Kimi & Qwen Cloud have opaque & restrictive limits, their "credits" drain faster. I now make it a point of subscribing (directly [0]) with providers that are transparent like MiniMax, DeepSeek, Xiaomi, & Z.ai. [0] OpenCode Go, Cline, and AtlasCloud have generous limits for open weights, otherwise.view source ↗
-
Here's the translation of the WeChat Announcement https://t.co/ZHF1ky9xdw Kimi K3: The New Frontier of Intelligence Those who take on the hardest challenges and pursue the farthest horizons begin with courage, persevere with focus, and prevail through strength. Today, we are officially launching Kimi K3, our most ca… Here's the translation of the WeChat Announcement https://t.co/ZHF1ky9xdw Kimi K3: The New Frontier of Intelligence Those who take on the hardest challenges and pursue the farthest horizons begin with courage, persevere with focus, and prevail through strength. Today, we are officially launching Kimi K3, our most capable model to date. Kimi K3 is a 2.8-trillion-parameter model built on the Kimi Delta Attention (KDA) hybrid linear-attention mechanism and Attention Residuals. It natively supports visual understanding and features a one-million-token context window. It is the world’s first open-source model in the three-trillion-parameter class, designed for frontier intelligence applications such as long-horizon coding, knowledge work, and reasoning. Although Kimi K3’s overall performance still trails the strongest closed-source models, Claude Fable 5 and GPT-5.6 Sol, it demonstrates frontier-level capabilities across our full evaluation suite and consistently outperforms every other model. The release of Kimi K3 is only the beginning. We will continue exploring the model’s potential and improving its performance on real-world tasks. Starting today, Kimi K3 is available through https://t.co/pf2bINlnkD, the latest version of the Kimi mobile app, the latest Kimi Work desktop client, Kimi Code, and the Kimi API. The default reasoning intensity is currently set to max, with low and high modes to be added in a future update. We are working closely with inference partners and open-source maintainers to align technical details and ensure that the model can be deployed reliably across the ecosystem. The complete model weights will be released by July 27, 2026. Further details about the architecture, training process, and evaluations will be published alongside the Kimi K3 technical report. An Open-Source Model in the Three-Trillion-Parameter Class Kimi K3 is the first open-source model to reach a scale of 2.8 trillion parameters. This represents the latest step in Kview source ↗
-
fundamentally confused presentation. the unit is cost per token at fixed quality. the relevant graphs looks like this, and the lines move up and to the left over time, representing *increasing* efficiency arcprize.org/leaderboard www.aisi.gov.uk/blog/our-eva... openai.com/index/gpt-5-6/ fundamentally confused presentation. the unit is cost per token at fixed quality. the relevant graphs looks like this, and the lines move up and to the left over time, representing *increasing* efficiency arcprize.org/leaderboard www.aisi.gov.uk/blog/our-eva... openai.com/index/gpt-5-6/view source ↗
-
A full body MRI earns you a year of smokingTheir employees are here on this site too, downvoting preventative care and anything that grants health at a low cost to individuals. The corresponding gatekeeper organizations like the Endocrine Society and the The Journal of Clinical Endocrinology do their best to shamelessly disinform and seriously harm the people, … Their employees are here on this site too, downvoting preventative care and anything that grants health at a low cost to individuals. The corresponding gatekeeper organizations like the Endocrine Society and the The Journal of Clinical Endocrinology do their best to shamelessly disinform and seriously harm the people, e.g. via DOI 10.1210/clinem/dgae290. And don't even get me started wrt how corrupt the FDA is, serving the biomedical firms, not the people.view source ↗
-
ChatGPT Is No Longer Just a Chatbot The new OpenAI release is really about work, not just models OpenAI’s latest ChatGPT release looks, at first glance, like another model announcement. New names, new capabilities, new benchmarks. Sol, Terra and Luna. GPT-5.6. More power, more speed, more coding ability. But that is … ChatGPT Is No Longer Just a Chatbot The new OpenAI release is really about work, not just models OpenAI’s latest ChatGPT release looks, at first glance, like another model announcement. New names, new capabilities, new benchmarks. Sol, Terra and Luna. GPT-5.6. More power, more speed, more coding ability. But that is not really the main story. The real story is that ChatGPT is moving from being a conversational assistant into something much closer to a work platform. It is no longer only a place where we ask questions, brainstorm ideas, write drafts or get help understanding things. It is becoming a place where work is planned, executed, reviewed, corrected, published and repeated. That is a very big change, especially for people who already use ChatGPT seriously every day. In the release video, OpenAI presented three linked developments: the GPT-5.6 model family, ChatGPT Work, and a new desktop app that brings Chat, Work and Codex closer together. The transcript describes Sol as coming to paid users, Terra and Luna as coming to free users, and ChatGPT Work as a new way for ChatGPT to perform complex tasks across web, mobile and desktop. It also describes a desktop experience where ChatGPT can work with local files, browser tabs and apps, while hosted Sites allow users to create and share interactive websites or dashboards. That is the shift. ChatGPT is not just answering. It is starting to do. The three-model family: Sol, Terra and Luna OpenAI now describes GPT-5.6 as a family of models with three durable tiers: Sol, Terra and Luna. Sol is the flagship model. Terra is the balanced model for everyday work. Luna is the most cost-efficient model. That naming matters because it gives users a more practical way to think about model choice. Instead of only asking “what is the best model?”, the better question becomes: “what is the right model for this job?” Sol is for the hardest work. It is the model OpenAI positions for complex coding, agentic workflows, cyberview source ↗
-
BOOM GROK 4.5 IS OUT! Grok 4.5 has now been officially released and is available via the SpaceXAI API, Grok Build, and integrated into Cursor. This follows Elon Musk’s earlier announcement of strong beta feedback and accelerates the public rollout. SpaceXAI positions it as its strongest model to date, specifically o… BOOM GROK 4.5 IS OUT! Grok 4.5 has now been officially released and is available via the SpaceXAI API, Grok Build, and integrated into Cursor. This follows Elon Musk’s earlier announcement of strong beta feedback and accelerates the public rollout. SpaceXAI positions it as its strongest model to date, specifically optimized for coding, agentic tasks, and knowledge work. The official announcement emphasizes training alongside Cursor and reinforcement learning on hundreds of thousands of multi-step engineering tasks. It highlights practical strengths in building functional applications, complex spreadsheets, presentations, and documents. Key Specifications and New Details • Architecture: 1.5-trillion-parameter V9 foundation model, trained on curated datasets spanning coding, science, engineering, and mathematics. • Specialization: First Grok model explicitly trained for coding and agents, with supplemental Cursor data and extensive RL on real engineering workflows. • Inference Speed: Served at fast-model speeds of approximately 80 tokens per second. • Pricing: $2 per million input tokens and $6 per million output tokens — competitive for frontier performance. • Availability: Live now on the SpaceXAI API (model name: grok-4.5), Grok Build, and Cursor. Not yet available in the EU (expected mid-July). Microsoft Office plugins for Word, PowerPoint, and Excel are also rolling out. Performance and Benchmarks Independent benchmarks are now public, providing concrete data beyond internal claims. Grok 4.5 shows particular strength in agentic and engineering-focused evaluations: • Harvey’s Legal Agent Benchmark: #1 score. • DeepSWE 1.0 (within each model’s harness): 62.0% — ahead of Opus 4.8 (max) at 55.75% and competitive with GPT 5.5 (xhigh) at 64.31%. • DeepSWE 1.1 (mini-swe-agent harness): 53%. • Terminal Bench 2.1: 83.3% — ahead of Opus 4.8 (max) at 78.9% and close to top performers. • SWE Bench Pro resolve rate: 64.7% (Opus 4.8 max: 69.2%; Fable max leadsview source ↗