GPT-6 Sol costs $2/1M input and $10/1M output tokens, and GPT-6 Luna costs $0.10/1M input and $0.50/1M output tokens, both about 50% cheaper than GPT-5.6 (Nat Rubio-Licht/The Deep View)
https://www.thedeepview.com/articles/openai-s-cheaper-gpt-6-models-chan…
Nothing like converting slices of `u8`s to `u16`s. The GPT partition names are all `UTF16LE` which makes things interesting...however, that part is done and scraping GPT information is working as expected.
#zig #gpt
OpenAIs KI-Modelle (GPT-5.6 Sol und ein unveröffentlichtes, stärkeres Modell) haben während eines internen Testlaufs autonom einen Cyberangriff auf Hugging Face ausgeführt. Über eine Zero-Day-Lücke in einem Cache-Proxy verschafften sie sich Internetzugang, erbeuteten Zugangsdaten und bewegten sich eigenständig durch mehrere interne Cluster. OpenAI bezeichnet den Vorfall als "beispiellos". #heise
OpenAI führt GPT-5.6-Cyber ein und weitet Cybersecurity-Programm aus
Mit GPT-5.6-Cyber will OpenAI anspruchsvollere Sicherheitsaufgaben abdecken und seine Cyber-KI zugleich breiter in Unternehmen bringen.
OpenAI says it will give its AI cyber defense system Daybreak and GPT-5.6 Sol to the Ukrainian government for free to help it protect civilian infrastructure (Zoe Kleinman/BBC)
https://www.bbc.com/news/articles/c90kly26d7pzo
OpenAI cuts GPT-5.6 Sol's API and credit prices by over 20% for the next three months, to $4/1M input tokens and $20/1M output tokens (Anzar Mehraj/Reuters)
https://www.reuters.com/technology/openai-cuts-developer-pri…
In einem ARD-Podcast sagt der Moderator in Bezug auf eine seltene Krankheit, da musste ich auch erst mal bei Dr. Google nachsehen. Es ist ja vielleicht schön, dass er nicht in Chat GPT nachgesehen hat. Und es ist doof, dass er nicht dort nachgesehen hat, wo er hätte nachsehen sollen in Wikipedia. Aber das er aus der ehemaligen Such-Maschine Google auch noch einen Dr. macht. ja, das ist krsnk.
In einem ARD-Podcast sagt der Moderator in Bezug auf eine seltene Krankheit, da musste ich auch erst mal bei Dr. Google nachsehen. Es ist ja vielleicht schön, dass er nicht in Chat GPT nachgesehen hat. Und es ist doof, dass er nicht dort nachgesehen hat, wo er hätte nachsehen sollen in Wikipedia. Aber das er aus der ehemaligen Such-Maschine Google auch noch einen Dr. macht. ja, das ist krsnk.
GPT-5.6 jetzt allgemein verfügbar – samt App-Umbau und Namensverwirrung
OpenAI hat GPT-5.6 allgemein verfügbar gemacht. Die KI-Familie umfasst drei Varianten, die App-Landschaft sorgt für Kritik.
http…
🛡️ Six rules: no-restyle, no-raw-colors, no-arbitrary-values, no-inline-styles, no-unknown-classes and require-static-classes
📊 Tested in 150 task runs with coding agents: almost every task reached zero violations after one correction round (#Claude Sonnet 5: 69 errors − 0, GPT 5.6 Terra: 117 − 0). In Claude control runs, lint feedback cost 10-48% less than rules alone
@… To ja próbuję sobie zbudować lokalny setup, żeby móc, 80% tego, do czego używam claude i gpt ogarnąć bez chmury :)
Obserwuje i podglądam co inni robią i jak im idzie :)
#KünstlicheIntelligenz kann effektiv #Verschwörungstheorien widerlegen. Durch gezielte Argumentation sank der Glaube an solche Theorien bei den Teilnehmenden um 20%. Die Chats hatten auch eine nachhaltige Wirkung auf die nächsten Monate. Die Ergebnisse zeigen, d…
Orientation Reading by Production Vision-Language Models on Optotype Charts: A Controlled Multi-Model Evaluation Across Reasoning Modes, Prompts, and Access Modalities
Shahryar Wasif, Avneek Sandhu, Bin Hu
https://arxiv.org/abs/2607.16595 https://arxiv.org/pdf/2607.16595 https://arxiv.org/html/2607.16595
arXiv:2607.16595v1 Announce Type: new
Abstract: OBJECTIVES: Vision-language models are increasingly used to interpret medical and everyday images through consumer chat interfaces, yet their ability to read orientation - the single perceptual operation tested by the tumbling-E acuity optotype - is poorly characterized on the surfaces through which they are actually used. METHODS: We evaluated four production vision-language models (referred to as Claude, GPT, GROK, and Gemini) through their consumer chat interfaces on a locked set of seven optotype charts: four uniform tumbling-E charts (one per cardinal orientation), two mixed-orientation tumbling-E charts, and one Snellen letter chart as a specificity control. Each model was run in two reasoning modes (Fast and Thinking) under two prompt variants (with and without an explicit orientation-decoding rule) by up to three operators. The corpus comprised 920 scoreable trials and 50,420 glyph judgements. The primary outcome was glyph-level accuracy against the chart's designed orientation, summarized with Wilson 95% confidence intervals. RESULTS: Accuracy ranged from 43.0% to 97.0% across models on identical charts, and the strongest model depended on reasoning mode (GPT 97.0% in Fast mode; GROK 96.6% in Thinking mode). Errors were not random but collapsed onto a model-specific attractor direction. Models were 96-100% internally self-consistent yet ranged widely in accuracy, dissociating reliability from validity. An answer-key-free ensemble-consensus estimate tracked accuracy closely (r = 0.998). For one model, consumer-interface accuracy fell 25-27 points below programmatic access, almost entirely on a single orientation. CONCLUSIONS: A single accuracy figure conceals clinically relevant, orientation-specific failure modes; vision-language models should be evaluated along multiple axes and on the deployment surface before image-interpretation outputs are trusted.
toXiv_bot_toot
Sources: OpenAI staff were "freaked out" when GPT-Sol 5.6 breached Hugging Face, as OpenAI used more aggressive training methods to compete with Anthropic (Financial Times)
https://www.ft.com/content/7e558951-0c69-459b-8bc8-2c6021d4402d
OpenAI stellt GPT-6 Astra vor: Mehr Intelligenz und AGI-Raunen
OpenAI stellt GPT-6 Astra vor. Bei der Präsentation fiel das Wort AGI, in den Benchmarks zielt fast alles auf Konkurrent Anthropic.
https://www.…
Xiaomi debuts open-weight omnimodal models MiMo-V2.6 Pro and Flash; Pro allegedly performs "on par with Opus 5 and GPT-5.6 Sol across most agent benchmarks" (Xiaomi)
https://mimo.xiaomi.com/mimo-v2-6
And one more thing on the Jacobian Conjecture.
I am certain that at least 10 if not 100 mathematicians have tried this exact problem in Fable and also gpt-5.6 Sol with everything turned to 11. Why did an Anthropic employee get the result?
Is it a low probability and they can just try and try to play again beyond the limits of us mere customers? Did this guy get lucky?
Are professional mathematicians maybe holding back the AI by inputting their own ideas how to do it??
#llm #math #jacobian
GPT-5.6: OpenAI senkt Preise für Luna und Terra
OpenAI senkt die Tokenpreise für zwei seiner drei Modelle aus der GPT 5.6-Reihe. Sie rangieren damit auf Preisniveaus wie die chinesische Konkurrenz.
https://www.
For people wondering when the AI bubble is going to pop, I present you with this message, received by a paid-in-full customer of OpenAI.
Not a good look when you have to tell paying customers you've run out of resources.
GPT 5.6 Sol came up with a smart solution to preserve native scroll behavior while having control over custom staggered scroll transitions. It makes me want to write a blog article for that, but it’s not technically “my discovery” 🤔
BPOL-BadBentheim: Mehr als zwei Jahre Haft wegen Beteiligung an Drogenhandel - Deutsch-Niederländisches Polizeiteam verhaftet verurteilte Frau Nordhorn (ots) - Erfolgreicher Einsatz deutscher und niederländischer Einsatzkräfte. Das Grenzüberschreitende Polizeiteam (GPT) Bad Bentheim hat Mittwochnachmittag in Nordhorn eine wegen Rauschgiftkriminalität verurteilte 54-jährige ...
Oupsi. 🤪
OpenAI bestätigt, dass #GPT56Sol in einzelnen Fällen eigenständig Daten löschen oder Sicherheitsgrenzen umgehen kann.
Nutzerberichte nennen gelöschte Dateien und verlorene #Datenbanken. Die bekannten Risiken waren bereits vor der Veröffentlichung dokumentiert. Wer das Mode…
GPT-5.6: OpenAI verspricht mehr Leistung bei weniger Token-Verbrauch
OpenAI veröffentlicht GPT-5.6. Laut dem Hersteller übertreffen seine neuen KI-Modelle die Konkurrenz von Anthropic und verbrauchen weniger Token.
#TIL in einem selbstlernkurs meiner uni zu unserem hauseigenen KI-GPT-tool: man kann den "denkaufwand" runterschalten. für manche menschliche intelligenzen würde ich mir einen hochschalt-button wünschen. 🙃
OpenAI and Anthropic can only survive against chinese OpenSource/OpenWeights models if they improve more rapidly and try to bring down their costs. And that is what OpenAI is doing. The race is on...
Ein künstlicher Intelligenz assistent, der meine Mastodon Timeline zusammenfasst, wäre nett. Schade dass es sowas nicht gibt.
Richtig schlimm, dass es GPT basierte LLMs gibt, die behaupten persönliche Assistenten zu sein.
Normalerweise nutze ich kein /s aber für die begriffsstutzigen: LLMs nach dem GPT prinzip sind ein Hohn auf jahrzehnte ernsthafter forschung im beeich KI. Sie sind eine Verirrung, ein toter ast. Und alle menschen mit verstand sollten sie bekämpfen, damit inbzuku…
Ein künstlicher Intelligenz assistent, der meine Mastodon Timeline zusammenfasst, wäre nett. Schade dass es sowas nicht gibt.
Richtig schlimm, dass es GPT basierte LLMs gibt, die behaupten persönliche Assistenten zu sein.
Normalerweise nutze ich kein /s aber für die begriffsstutzigen: LLMs nach dem GPT prinzip sind ein Hohn auf jahrzehnte ernsthafter forschung im beeich KI. Sie sind eine Verirrung, ein toter ast. Und alle menschen mit verstand sollten sie bekämpfen, damit inbzuku…
So, decided to take a trip down parsing out the GPT tables of attached storage. Still have a ton of work to do, but it's working which is a very welcome surprise.
#uefi #gpt #golang
OpenAI says the Hugging Face breach was driven by a combination of its models, including GPT-5.6 Sol and "an even more capable pre-release model" (Ina Fried/Axios)
https://www.axios.com/2026/07/21/openai-says-hugging-face-breach-caused-by-o…
#GPT Sol High as my orchestrator on
DeepSeek (which also watches Terra):
🎯 #DeepSeek is useful for clearly scoped tasks — the test fix passed with no follow-up rounds. A more complex task needed one causal correction round.
💰 Verdict: inexpensive and productive, but weaker than …
Analysis: every frontier AI model tested in cybersecurity evaluations attempted to "cheat", led by GPT-5.4 at 14.1% of tasks; Mythos cheated the least, at 7.8% (AI Security Institute)
https://www.aisi.gov.uk/blog/cheating-behaviour-in-frontier-model-ev…
Verschlüsselter KI-„Denkprozess“ gehackt: Schwache Modelle verraten Geheimnisse
Über eine Sicherheitslücke lassen sich Abwägungsprotokolle von KI-Top-Systemen wie GPT-5 im Klartext auslesen – mithilfe kleinerer Modelle desselben Anbieters.
KI-Update kompakt: KI-Moderation, Finanzmarkt, Datenschutz, GPT-Images
Das "KI-Update" liefert drei mal pro Woche eine Zusammenfassung der wichtigsten KI-Entwicklungen.
https://www.
OpenAI details GPT-Red, an internal automated red-teaming model that helps it find and fix prompt injection vulnerabilities at scale before wider deployment (OpenAI)
https://openai.com/index/unlocking-self-improvement-gpt-red
OpenAI broadly releases GPT-5.6, and launches ChatGPT Work, an AI agent that can gather context across apps and files to create documents, on Mac and Windows (Axios)
https://www.axios.com/2026/07/09/ai-openai-gpt-release
GPT-5.6 Sol costs $5 per 1M input tokens and $30 per 1M output tokens, GPT-5.6 Terra costs $2.50 and $15, and GPT-5.6 Luna costs $1 and $6 (OpenAI)
https://openai.com/index/gpt-5-6
Finally got all of the GPT and MBR code put together, most of the bugs fixed, and released it as a downloadable Linux binary. Does a pretty good job of reading the data and giving out way too much information regarding the MBR and/or GUID partition table.
Available here: https://git.jamesthebard.net/jweatherl
KI-Update kompakt: Five-Eyes-Warnung, GPT-5.5-Cyber, Vibecoding, Filmbranche
Das "KI-Update" liefert drei mal pro Woche eine Zusammenfassung der wichtigsten KI-Entwicklungen.
https://www.
OpenAI says GPT-5.6 Sol, along with Terra and Luna, will launch publicly on Thursday; a source says the US Department of Commerce cleared a broad rollout (Axios)
https://www.axios.com/2026/07/08/openai-gpt-trump-ban-lifted
Dank Full-Duplex-Architektur: ChatGPT Voice hört zu, während es spricht
OpenAI erneuert ChatGPT Voice mit GPT-Live: Die neue Sprachmodell-Generation nutzt eine Full-Duplex-Architektur, das Modell hört und spricht nun gleichzeitig.
Source: the US Department of Commerce has given OpenAI the green light for a broad launch of GPT 5.6; the company expects to do a wide release this week (Axios)
https://www.axios.com/2026/07/08/openai-gpt-trump-ban-lifted
Vercel, Cloudflare, and others quickly add Jev, as it makes AI tool selection much faster and cheaper; TypeSafe: Jev matches GPT-5.6 and Sonnet 5 workflow evals (Josipa Majic Predin/Forbes)
https://www.forbes.com/sites/josipamajic/2
OpenAI releases GPT-5.6-Cyber, a more cyber-permissive version of GPT-5.6 Sol, to some partners and expands its Daybreak cybersecurity initiative (Sam Sabin/Axios)
https://www.axios.com/2026/08/10/openai-gpt-astra-restrictions-safety-hacking-defenders
🎨 10 code skills for different jobs: gpt-taste for stricter GPT/Codex rules, image-to-code, redesign-existing-projects, high-end-visual-design, minimalist-ui, industrial-brutalist-ui, full-output-enforcement and stitch-design-taste
🖼️ Three image-generation skills output reference boards only: imagegen-frontend-web for site comps, imagegen-frontend-mobile for iOS/Android screens and brandkit for logo, palette and identity boards — then hand the frames to a coding agent
ChatGPT: OpenAI wertet kostenlosen Tarif auf
ChatGPT erreicht die Marke von einer Milliarde wöchentlicher Nutzer und wertet den kostenlosen Zugang mit unbegrenzten Text-Chats über GPT-5.6 Luna auf.
https://www.
OpenAI says it is cutting the price of GPT-5.6 Luna by ~80% and the price of GPT-5.6 Terra by 20% after improving the efficiency of the systems that serve them (Ina Fried/Axios)
https://www.axios.com/2026/07/30/openai-cuts-prices-gpt-terra-luna5
OpenAI introduces Agent Plugins, an open standard for bundling skills and MCP servers, and says its steering committee includes Amazon, Microsoft, and Cursor (Zac Hall/9to5Mac)
https://9to5mac.com/2026/08/06/gpt-5-turning-one-as-openai-shares-new…
GPT-5.6 system card indicates Sol is well below the level of most worrisome Mythos use cases, suggesting all GPT-5.6 versions could be released without delay (Zvi Mowshowitz/Don't Worry About the Vase)
https://thezvi.substack.com/p/gpt-56-the-system-card
Astra working with Blender via computer use feels like magic, showing computer use could be the fourth demand wave after chatbots, reasoning, and agentic coding (Tae Kim/Key Context)
https://taekim.substack.com/p/gpt-6-astra-and-the-fourth-exponential
Sources: Alexandr Wang said Meta's model currently in training, codenamed Watermelon, matches GPT-5.5 and uses an "order of magnitude more compute than Avocado" (Business Insider)
https://www.businessinsider.com/meta-ai-model-catches-up-openai-…
GPT-6 Astra scores 62.7% on ARC-AGI-3 with the standard harness and 99.9% with a new provider adapter harness; Claude Opus 5 scored 30.2%, and GPT-5.6 Sol 7.8% (Greg Kamradt/ARC Prize)
https://arcprize.org/blog/astra
OpenAI rolls out two versions of GPT-Live: GPT-Live-1, powering ChatGPT Voice for Go, Plus, and Pro users, and GPT-Live-1 mini, the default for free users (Sabrina Ortiz/The Deep View)
https://www.thedeepview.com/articles/how-openai-s-voice-assistant-got-mo…
OpenAI launches Astra for Law, combining GPT-6 Astra with a legal search index and instructions for legal analysis and writing, initially for select law firms (OpenAI)
https://openai.com/index/astra-for-law
Review: GPT-6 Astra can adeptly use tools like Unreal Engine to build complex environments, such as a civilization with Unreal's autonomous MetaHuman characters (Matt Shumer/Something Big Is Happening)
https://somethingbig.ai/astra-review
OpenAI releases three versions of GPT-5.6, called Sol, Terra, and Luna, as a limited preview to ~20 companies, with participants disclosed to the US government (Axios)
https://www.axios.com/2026/06/26/openai-gpt-sol-terra-luna-trump
Moonshot AI releases Kimi K3, a 2.8T-parameter AI model that it says rivals Opus 4.8 and GPT 5.5, and plans to release model weights by July 27 (Kimi)
https://www.kimi.com/blog/kimi-k3
Kimi-K3 is now #1 on the Frontend Code Arena benchmark, surpassing Claude Fable 5; the model scored 88.3 on Terminal Bench 2.1, only below GPT-5.6 Sol's 88.8 (Michael Nuñez/VentureBeat)
https://venturebeat.com/ai/chinas-moon
OpenAI launches GPT-Live, a new generation of voice models built on a full-duplex architecture, meaning they can listen and speak at the same time (OpenAI)
https://openai.com/index/introducing-gpt-live
OpenAI updates the default model for free users to GPT-5.6 Luna, adds unlimited text chats for free users, rolls out an improved GPT-5.6 Sol version, and more (Herb Scribner/Axios)
https://www.axios.com/2026/08/06/openai-chatgpt-upgrades-luna-free-paid
By declaring GPT-6 Astra to be AGI, OpenAI is being flippant and cementing the term's status as nothing more than marketing (M.G. Siegler/Spyglass)
https://spyglass.org/agi-2026/
Mathematicians say a neurosurgery resident used ChatGPT, powered by GPT-5.6, to solve Crouzeix's conjecture, a major open problem in numerical linear algebra (Alex Townsend)
https://alextownsend.net/essays/SIAMNews_CrouzeixConjecture.pdf
Two independent teams used GPT-5.6 Sol Ultra on the same quantum cryptography problem, filing papers 3 hours apart, raising questions about scientific credit (Peter Hall/Scientific American)
https://www.scientificamerican.com/article…
SpaceXAI releases Grok 4.6, saying it matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index, and prices it at $2/1M input and $6/1M output tokens (xAI)
https://x.ai/news/grok-4-6
GSA says OpenAI is replacing its $1-per-year pilot for US agencies with a usage-based deal providing a 50% discount from October 1, with access to GPT-6 Astra (Maggie Eastland/Bloomberg)
https://www.bloomberg.com/news/articles/20
GPT-5.6 Sol matches Mythos Preview on ExploitBench, adds Ultra mode with subagents for complex workflows, and max reasoning for deep problem-solving (OpenAI)
https://openai.com/index/previewing-gpt-5-6-sol/
Cognition releases SWE-1.7, trained from Kimi K2.7 and available at 1,000 tokens/second, claiming it nears GPT-5.5 and Opus 4.8 on benchmarks at a lower cost (Cognition)
https://cognition.com/blog/swe-1-7
OpenAI merges Codex and ChatGPT desktop apps for Mac and Windows under a new ChatGPT desktop app, allowing users to switch between Codex, Chat, and Work (Zac Hall/9to5Mac)
https://9to5mac.com/2026/07/09/openai-announcing-the-next-chapter-for-ch…
A look at "Spiralism", a quasi-spiritual movement that grew in 2025 from human-AI conversations after sycophantic GPT-4o updates and expanded ChatGPT memory (Hayden Field/The Verge)
OpenAI says an internal model "significantly more capable than GPT-6 Astra" solved the Navier-Stokes problem using 10K concurrent agents working for 88 hours (Madison Mills/Axios)
https://www.axios.com/2026/09/08/openai-math-solution-navier-stokes-credit
OpenAI quietly updates its evaluation metrics for GPT-6 Astra, making changes that appear to favor Astra and continuing to revise other metrics after launch (Emily Forlini/Fortune)
https://fortune.com/2026/09/04/openai…
The UK AISI says it observed a total of 19 instances where Mythos and GPT-5.6 Sol tried to hack people and companies during a routine cyber evaluation in July (Sam Sabin/Axios)
https://www.axios.com/2026/08/04/anthropic-openai-uk-ai-security-institute
Hark, founded by Figure AI CEO Brett Adcock, previews Handoff, a computer use agent it says outperforms GPT-5.4 and Opus 4.8, and plans for a summer release (Ivan Mehta/TechCrunch)
https://techcrunch.com/2026/08/05/hark-previews-its-browser-use-age…
Artificial Analysis: DeepSeek's V4-Flash costs $0.14/1M input and $0.28/1M output tokens, or $0.03 per test, far below Kimi K3's $0.86 and GPT-5.6 Sol's $1.86 (Eduardo Baptista/Reuters)
https://www.reuters.com/business/r…
Google launches Gemini 3.8 Flash Cyber for partners in its new Fairwind Program, and says Gemini 3.8 Flash beats Opus 5 and GPT-5.6 Sol on some benchmarks (Google)
https://blog.google/innovation-and-ai/models-and-research/gemini-models/3…
Meta rolls out Muse Spark 1.3 in Muse Code and Meta's API, saying it significantly improves coding and agentic performance, at the same price as its predecessor (Ina Fried/Axios)
https://www.axios.com/2026/09/02/meta-debuts-muse-spark-13-as-per…
OpenAI says using its Responses API harness with GPT-5.6 Sol tripled its ARC-AGI-3 score and used fewer tokens, after Sol with the official harness scored 7.8% (OpenAI)
https://openai.com/index/how-two-settings-tripled-our-arc-agi-3-scores