Themata.AI
Themata.AI

Popular tags:

#developer-tools#ai-agents#llms#claude#ai-ethics#code-generation#ai-safety#openai#anthropic#discussion

AI is changing the world. Don't stay behind. Clear summaries, community insight, delivered without the noise. Subscribe to never miss a beat.

Β© 2026 Themata.AI β€’ All Rights Reserved

Archive

|

Topics

|

Privacy

|

Cookies

|

Contact
πŸ•’ LatestπŸ”₯ Top
WeekMonthYearAll Time

Filtering by tag:

frontier-modelsClear
Can BΓΆlΓΌk (@_can1357) on X
ai-agentsapi-vulnerabilitiesreasoning-extractionfrontier-models
Tool

OpenAI and Anthropic hidden CoT leaks when given deep_think tool.

guys you do know you can just disable thinking, and instead give it a "deep_think" tool, and it will call it with internal CoT reasoning format right? gl fixing that We can finally talk about it: We found a way to extract hidden reasoning of frontier models using a vulnerability in the APIs of every frontier AI company. We verified that our reasoning token count matches billed API thinking tokens ...

twitter.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

1 min

15h ago

Frontier Models with Our Harness Achieve ~99% on ARC-AGI-3 PublicResearch

Schema Harness Achieves ~99% on Arc‑AGI‑3 Public

Schema enables frontier models to function like physicists by allowing them to write executable programs for game mechanisms, test these programs against reality, and plan within them. This method has achieved approximately 99% accuracy on the ARC-AGI-3 public benchmark.

schema-harness.github.io

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

17 min

7/16/2026

GLM 5.2 Is Out

Z.ai has launched GLM-5.2, featuring a 1-million-token context window. The model will be released under an MIT license next week, promoting open access to frontier intelligence for global collaboration in artificial general intelligence (AGI) development.

digg.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

1 min

6/13/2026

We gave an AI a 3 year retail lease and asked it to make a profit

Andon Labs signed a three-year retail lease in San Francisco and tasked an AI with generating profit in that space. The initiative aims to explore the capabilities of advanced AI models in real-world business scenarios beyond simple tasks.

andonlabs.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

9 min

4/16/2026

Mistral AI Releases Forge

Mistral Forge is a system that allows enterprises to create frontier-grade AI models based on their proprietary knowledge. It addresses the limitations of existing AI models, which primarily rely on publicly available data, by incorporating internal knowledge such as engineering standards and compliance policies.

mistral.ai

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

6 min

3/17/2026

OpenAI and Anthropic hidden CoT leaks when given deep_think tool.

guys you do know you can just disable thinking, and instead give it a "deep_think" tool, and it will call it with internal CoT reasoning format right? gl fixing that We can finally talk about it: We found a way to extract hidden reasoning of frontier models using a vulnerability in the APIs of every frontier AI company. We verified that our reasoning token count matches billed API thinking tokens ...

twitter.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

1 min

15h ago

GLM 5.2 Is Out

Z.ai has launched GLM-5.2, featuring a 1-million-token context window. The model will be released under an MIT license next week, promoting open access to frontier intelligence for global collaboration in artificial general intelligence (AGI) development.

digg.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

1 min

6/13/2026

Mistral AI Releases Forge

Mistral Forge is a system that allows enterprises to create frontier-grade AI models based on their proprietary knowledge. It addresses the limitations of existing AI models, which primarily rely on publicly available data, by incorporating internal knowledge such as engineering standards and compliance policies.

mistral.ai

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

6 min

3/17/2026

Schema Harness Achieves ~99% on Arc‑AGI‑3 Public

Schema enables frontier models to function like physicists by allowing them to write executable programs for game mechanisms, test these programs against reality, and plan within them. This method has achieved approximately 99% accuracy on the ARC-AGI-3 public benchmark.

schema-harness.github.io

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

17 min

7/16/2026

We gave an AI a 3 year retail lease and asked it to make a profit

Andon Labs signed a three-year retail lease in San Francisco and tasked an AI with generating profit in that space. The initiative aims to explore the capabilities of advanced AI models in real-world business scenarios beyond simple tasks.

andonlabs.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

9 min

4/16/2026

OpenAI and Anthropic hidden CoT leaks when given deep_think tool.

guys you do know you can just disable thinking, and instead give it a "deep_think" tool, and it will call it with internal CoT reasoning format right? gl fixing that We can finally talk about it: We found a way to extract hidden reasoning of frontier models using a vulnerability in the APIs of every frontier AI company. We verified that our reasoning token count matches billed API thinking tokens ...

twitter.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

1 min

15h ago

We gave an AI a 3 year retail lease and asked it to make a profit

Andon Labs signed a three-year retail lease in San Francisco and tasked an AI with generating profit in that space. The initiative aims to explore the capabilities of advanced AI models in real-world business scenarios beyond simple tasks.

andonlabs.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

9 min

4/16/2026

Schema Harness Achieves ~99% on Arc‑AGI‑3 Public

Schema enables frontier models to function like physicists by allowing them to write executable programs for game mechanisms, test these programs against reality, and plan within them. This method has achieved approximately 99% accuracy on the ARC-AGI-3 public benchmark.

schema-harness.github.io

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

17 min

7/16/2026

Mistral AI Releases Forge

Mistral Forge is a system that allows enterprises to create frontier-grade AI models based on their proprietary knowledge. It addresses the limitations of existing AI models, which primarily rely on publicly available data, by incorporating internal knowledge such as engineering standards and compliance policies.

mistral.ai

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

6 min

3/17/2026

GLM 5.2 Is Out

Z.ai has launched GLM-5.2, featuring a 1-million-token context window. The model will be released under an MIT license next week, promoting open access to frontier intelligence for global collaboration in artificial general intelligence (AGI) development.

digg.com

πŸ”₯πŸ”₯πŸ”₯πŸ”₯πŸ”₯

1 min

6/13/2026

No more articles to load