AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,499 results
14 Jul 2026

Exactly one year ago was the craziest 72 hours I’d experience in Silicon Valley. I was thrust into a situation where many members of the Win…

Model ReleasesDGX agent

Exactly one year ago was the craziest 72 hours I’d experience in Silicon Valley. I was thrust into a situation where many members of the Windsurf team had moved on to Google, recruiters were reaching

hello!

Model ReleasesDGX agent

hello! Hello. We have reached 8M active users across Codex and ChatGPT Work. We are once again resetting the usage limits for all. And we continue to not have the 5h rate limit as well, allowing every

I want to have this in writing so I can refer back to this tweet before it inevitably becomes a consensus view on twitter and say 'I told yo…

Model ReleasesDGX agent

I want to have this in writing so I can refer back to this tweet before it inevitably becomes a consensus view on twitter and say 'I told you so!' - The mass majority of researchers and academics in d

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

If you are use the Claude everything app, you pick between Home and Code. If you pick Home you get to pick between Chat & Cowork If you use …

Model ReleasesDGX agent

If you are use the Claude everything app, you pick between Home and Code. If you pick Home you get to pick between Chat & Cowork If you use the OpenAI everything app, you pick between ChatGPT Work & C

I'll open source this if it's interesting! But here's my first public artifact, a breakdown on my Mega Sceptile team: https://claude.ai/code…

Model ReleasesDGX agent

Thariq (@trq212) released his first public artifact on Claude.ai, a breakdown of his Mega Sceptile team titled “Mega Sceptile — Champions Field Guide.” The guide covers the team’s build, game plan, de

🔗 Official Ollama extension https://marketplace.visualstudio.com/items?itemName=Ollama.ollama

Local AiDGX agent

**Official Ollama Extension for VS Code** The extension on the Visual Studio Marketplace provides a dedicated Ollama language‑model integration for Visual Studio Code, enabling developers to run and a

Scaling UX testing with Amazon Nova Act: A new approach to user flow analysis

TutorialsDGX agent

Using generative AI enables parallel execution of comprehensive user flow testing at scale. This solution demonstrates how to build a cloud-deployed UX testing platform that automatically generates te

simonw/pedalican

Model ReleasesDGX agent

simonw/pedalican Clearly I wasn't paying attention when these were first announced back in May, but today I accidentally activated a 'pet' in Codex Desktop - a little animated robot, reminiscent of Cl

We’re open sourcing WANDR. WANDR is an internal benchmark we built and used for building deep and wide research capabilities inside Perplexi…

Model ReleasesDGX agent

We’re open sourcing WANDR. WANDR is an internal benchmark we built and used for building deep and wide research capabilities inside Perplexity Computer. https://research.perplexity.ai/articles/wandr-b

You can now use Claude inside After Effects. Higgsfield's new MCP connector lets Claude work inside your actual AE project. It can build com…

Model ReleasesDGX agent

You can now use Claude inside After Effects. Higgsfield's new MCP connector lets Claude work inside your actual AE project. It can build compositions, set keyframes, write expressions, and run the rep

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk …

Model ReleasesDGX agent

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk to anything with a phone number. Just one prompt. 15 seconds

13 Jul 2026

Computer use in Codex got very good on PC. Asking it to do something on your computer and having the cursor move under the control of a ghos…

Model ReleasesDGX agent

Computer use in Codex got very good on PC. Asking it to do something on your computer and having the cursor move under the control of a ghost is one of the things that makes you viscerally realize how

Creating my own interfaces in real-time to do the stuff I want to do is ridiculously empowering. Here's an example: - took ~10 app screensho…

Model ReleasesDGX agent

Creating my own interfaces in real-time to do the stuff I want to do is ridiculously empowering. Here's an example: - took ~10 app screenshots of things I wanted to fix - asked Claude to make a feedba

DOOMQL

Model ReleasesDGX agent

DOOMQL Peter Gostev built this using GPT-5.6 Sol. This is a lot of fun: DOOMQL started with a deliberately unreasonable question: what if SQLite were the game engine, not merely the place where a game

🆕 In Code They Act, In Proof We Trust — Erik Meijer last year, @solomonstre defined agents as 'an LLM that's wrecking its environment in a …

Model ReleasesDGX agent

🆕 In Code They Act, In Proof We Trust — Erik Meijer last year, @solomonstre defined agents as 'an LLM that's wrecking its environment in a loop', and @simonw coined the Lethal Trifecta for agents, tha

Launching UI for generative AI inference recommendations in Amazon SageMaker AI

Model ReleasesDGX agent

In this post, we introduce the UI for optimized generative AI inference recommendations in Amazon SageMaker AI Studio, a low-code no-code (LCNC) experience. The API already gives you programmatic acce

Livestream Alert: Run ComfyUI From Claude/Cursor with Comfy MCP Host: @PurzBeats Comfy MCP lets Claude, Cursor, Amp and almost any AI agent …

Model ReleasesDGX agent

Livestream Alert: Run ComfyUI From Claude/Cursor with Comfy MCP Host: @PurzBeats Comfy MCP lets Claude, Cursor, Amp and almost any AI agent you're already using build, run, and iterate real Comfy Clou

love to see two port cos collaborating😁 get private market data for your agents from @akta_pro with your @monid_ai account!

Model ReleasesDGX agent

love to see two port cos collaborating😁 get private market data for your agents from @akta_pro with your @monid_ai account! We just killed PitchBook. Introducing Claude for private market data. Your a

NEW: Grok 4.5 now scores highest on the SWE-Atlas-QnA benchmark, edging Claude Fable 5 & GPT-5.6 Sol.

Model ReleasesDGX agent

Grok 4.5 achieved the highest score on the SWE‑Atlas‑QnA benchmark, surpassing Claude Fable 5 and GPT‑5.6 Sol, according to a Polymarket tweet posted on 13 July 2026. The update highlights Grok 4.5’s

Standard RL benchmarks are episodic and stationary, so they don't capture the the characteristics of real-world deployment. Morpheus is a ne…

Model ReleasesDGX agent

Standard RL benchmarks are episodic and stationary, so they don't capture the the characteristics of real-world deployment. Morpheus is a new benchmark for continual learning that provides persistent

We've open sourced Ideogram V4 instant and fast, find them here! Fast: https://huggingface.co/fal/ideogram-v4-fast Instant: https://huggingf…

IndustryDGX agent

We've open sourced Ideogram V4 instant and fast, find them here! Fast: https://huggingface.co/fal/ideogram-v4-fast Instant: https://huggingface.co/fal/ideogram-v4-instant Try them out for free on fal

12 Jul 2026

I don't think Anthropic realizes how disruptive these changes are to users. I appreciate the extension, but please stop playing games. Eithe…

Model ReleasesDGX agent

I don't think Anthropic realizes how disruptive these changes are to users. I appreciate the extension, but please stop playing games. Either keep it under the subscriptions or put it under the API al

11 Jul 2026

ChatGPT still has study mode, but rather than /study you now have to type @ study It makes the AI act more like a tutor than a helpful assis…

Model ReleasesDGX agent

ChatGPT still has study mode, but rather than /study you now have to type @ study It makes the AI act more like a tutor than a helpful assistant, and some work suggests it is better if you are trying

GPT-Live is now fully rolled out to all ChatGPT users globally. We're also doubling everyone's voice usage limit for the whole weekend so yo…

Model ReleasesDGX agent

OpenAI announced the global rollout of GPT-Live to all ChatGPT users, enabling real-time interaction capabilities. As part of the announcement, the company temporarily doubled voice usage limits for a

Grok places second after Fable on real-world software engineering

Model ReleasesDGX agent

Grok places second after Fable on real-world software engineering Grok 4.5 from @SpaceXAI places #2 on the APEX-SWE leaderboard at 51.2% Pass@1 (±6.0), behind Fable 5 (65.5% ±6.2) on our benchmark for

I gave Fable the code: 'take this game and do something incredible with it to make it something very different. Be creative' It created DEEP…

Model ReleasesDGX agent

I gave Fable the code: 'take this game and do something incredible with it to make it something very different. Be creative' It created DEEP TIME: create a city, watch it be abandoned and forgotten, a

Why type when you can call? docs​.together​.ai​/call

ToolsDGX agent

Together AI's 'Why type when you can call?' post likely promotes their voice calling or audio interface capabilities as an alternative to text-based interactions with their AI models. The content prob

Wordle 1,847 4/6 ⬛⬛🟨⬛⬛ 🟨⬛🟩⬛⬛ ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This appears to be a Wordle game result post from Anthropic's X (Twitter) account, showing that the word was solved in 4 attempts on puzzle #1,847 using the standard color-coded feedback system (gray

10 Jul 2026

A law of robustness for two-layer neural networks with arbitrary weights

Model ReleasesDGX agent

arXiv:2607.07778v1 Announce Type: new Abstract: Bubeck, Li and Nagaraj conjectured that, for generic data, any two-layer neural network with m neurons that fits n noisy labels must have Lipschitz cons

Alignment Plausibility: A New Standard for Assuring AI in Healthcare

SafetyDGX agent

arXiv:2607.07766v1 Announce Type: new Abstract: Large language models (LLMs) have become significant providers of mental health support, yet they remain products of an attention economy whose operatio

Anyone can visit the Starship factory and launch site in Texas, as it is right next to on the public highway. It is incredibly inspiring to …

Model ReleasesDGX agent

Anyone can visit the Starship factory and launch site in Texas, as it is right next to on the public highway. It is incredibly inspiring to see! SpaceX has released the next episode of its new docuser

AutoAnchor: Stable Diffusion Unlearning Using Cross-Attention as a Manifold Surrogate

ResearchDGX agent

arXiv:2607.08337v1 Announce Type: new Abstract: Diffusion unlearning is essential for mitigating the generation of harmful or copyrighted content in text-to-image models. Current diffusion unlearning

Beware What You Autocomplete: Forensic Attribution of Backdoored Code Completions

ResearchDGX agent

arXiv:2607.08011v1 Announce Type: cross Abstract: Large language models have enabled powerful code completion systems that assist developers by predicting subsequent lines of code. However, these mode

Beyond Thermal Imaging: Inferring Thermophysical Properties from Time-Resolved Thermal Observations

Model ReleasesDGX agent

arXiv:2607.07962v1 Announce Type: cross Abstract: Inferring latent physical properties from sensory observations is a fundamental challenge in machine perception. Among available sensing modalities, t

BiasBench: A reproducible benchmark for tuning the biases of event cameras

Model ReleasesDGX agent

arXiv:2504.18235v2 Announce Type: replace Abstract: Event-based cameras are bio-inspired sensors that detect light changes asynchronously for each pixel. They are increasingly used in fields like comp

Bridging Modal Isolation in Interleaved Thinking: Supervising Modality Transitions via Stepwise Reinforcement

ResearchDGX agent

arXiv:2606.12886v2 Announce Type: replace-cross Abstract: Interleaved thinking, where a unified multimodal model alternates between textual reasoning and visual generation, has shown promise on spatia

Certified Interventional Fidelity: Anytime-Valid, Adaptive Evaluation of Causal Claims in Mechanistic Interpretability

ResearchDGX agent

arXiv:2607.08349v1 Announce Type: new Abstract: Mechanistic interpretability often evaluates explanations by intervening on a model: swapping hidden states, patching activations, ablating components,

Classical versus Deep Mirror-Symmetry Scoring: A Benchmark of Thirteen Methods

Model ReleasesDGX agent

arXiv:2607.08379v1 Announce Type: new Abstract: Quantifying how mirror-symmetric an image is about a given axis (symmetry scoring) underpins applications from visual aesthetics to medical imaging, yet

Claude Code on desktop now has an in-app browser. Claude can pull up docs, designs, or any other site. It can read, click through, and inter…

Model ReleasesDGX agent

Claude Code on desktop now has an in-app browser. Claude can pull up docs, designs, or any other site. It can read, click through, and interact the same way it does with your local dev servers. It's s

Conversational AI for Rapid Scientific Prototyping: A Case Study on ESA's ELOPE Competition

ApplicationsDGX agent

arXiv:2601.04920v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as coding partners, yet their role in accelerating scientific discovery remains underexplored. Th

Detecting Ladder Logic Bombs in IEC 61131-3 PLC Programs using ESBMC-PLC+: A Formal Verification Approach with Trigger Synthesis

SafetyDGX agent

arXiv:2607.08417v1 Announce Type: new Abstract: A Ladder Logic Bomb (LLB) is malicious control logic in a Programmable Logic Controller (PLC) program that lies dormant until a trigger activates a payl

DexVerse: A Modular Benchmark for Multi-Task, Multi-Embodiment Dexterous Manipulation

Model ReleasesDGX agent

arXiv:2607.08751v1 Announce Type: new Abstract: Building general-purpose dexterous manipulation policies requires benchmarks that go beyond isolated tasks to systematically evaluate policies across di

Equivariant Quantum Clustering with Differential Privacy: Parameter-Efficient Privacy-Preserving Analysis Across Heterogeneous Sensitive Datasets

Model ReleasesDGX agent

arXiv:2607.08092v1 Announce Type: cross Abstract: Privacy-preserving clustering is critical for analyzing sensitive data in healthcare, cybersecurity, and enterprise applications, where maintaining da

folks who’ve tried Claude Design — how was the experience? what works well and what could be better?

Model ReleasesDGX agent

Claude Design appears to be a design-focused tool or feature within Claude that users can interact with; this post solicits feedback from people who have tried it, asking them to evaluate what aspects

Formal Mechanisms for Market Stability in Self-Interested Agent Societies: A Marketplace Simulation Study

Model ReleasesDGX agent

arXiv:2607.08652v1 Announce Type: new Abstract: Self-interested agents, left unconstrained, tend toward defection in repeated social dilemmas, causing cooperative gains from trade to collapse. This pa

great viz by @BraceSproul to put this in perspective. the really important part of this is that OpenWiki is additive!

Model ReleasesDGX agent

great viz by @BraceSproul to put this in perspective. the really important part of this is that OpenWiki is additive! OpenWiki general purpose memory is meant to be complementary to codex/claude code

Grok Build

Model ReleasesDGX agent

Grok Build Grok 4.5 with Grok Build just ranked #1 on the SWE-Atlas-QnA benchmark with a score of 84 That puts it level with GPT-5.6 (max) Codex and ahead of Claude Code Fable 5 (max), Opus 4.8 (max),

How Deutsche Telekom is rewiring telecommunications with AI

Model ReleasesDGX agent

Deutsche Telekom is implementing AI technologies to modernize and optimize its telecommunications infrastructure and operations. The article likely discusses specific applications of AI in areas such

How to get 100% consistent product ads from one seedance 2.0 generation all directed from one chat with the hashtag Comfy MCP. @hellorob did…

Model ReleasesDGX agent

How to get 100% consistent product ads from one seedance 2.0 generation all directed from one chat with the hashtag Comfy MCP. @hellorob didn't let the agent improvise a pipeline. He pointed it at his

I think it kind of makes most sense if we think of it as training vs inference-scaling (while keeping in mind that Sol, Terra, Luna are also…

ResearchDGX agent

This post discusses the conceptual distinction between training-scaling and inference-scaling in machine learning models, while acknowledging related frameworks or concepts (possibly referencing Sol,

Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation

Model ReleasesDGX agent

arXiv:2607.08758v1 Announce Type: new Abstract: Scientific ideas rarely start from a blank page. They inherit mechanisms, repair known limitations, and recombine pieces of earlier work, much like biol

Key to this strategy is deterrence, 'Mutually Assured Compute Destruction' which gets its own section. It doesn't mention the generalization…

Model ReleasesDGX agent

Key to this strategy is deterrence, 'Mutually Assured Compute Destruction' which gets its own section. It doesn't mention the generalization Mutually Assured AI Malfunction (MAIM) from Schmidt, Wang,

Less Data, Faster Convergence: Goal-Driven Data Optimization for Multimodal Instruction Tuning

Model ReleasesDGX agent

arXiv:2603.12478v2 Announce Type: replace Abstract: Multimodal instruction tuning is often compute-inefficient because training budgets are spread across large mixed image-video pools whose utility is

Less Is More: Reducing Token Counts Without Compromising Performance

ResearchDGX agent

arXiv:2506.15138v2 Announce Type: replace-cross Abstract: Tokenization directly affects the inference efficiency of large language models, since fragmented tokenization increases sequence length and g

Like the Claude app it imitates, the new ChatGPT 'Super App', which merges Codex with ChatGPT, is a tangle of toggles and strange UI decisions (M.G. Siegler/Spyglass)

Model ReleasesDGX agent

M.G. Siegler / Spyglass: Like the Claude app it imitates, the new ChatGPT “Super App”, which merges Codex with ChatGPT, is a tangle of toggles and strange UI decisions — Even if it's ultimately the ri

Mixture of Enhanced-View Experts for Multi-Query Vehicle ReID and A Large-Scale Benchmark

Model ReleasesDGX agent

arXiv:2607.08085v1 Announce Type: new Abstract: Multi-query vehicle ReID aims to leverage complementary information from diverse views for robust feature learning. However, current methods suffer from

Modular Pretraining Enables Access Control

ResearchDGX agent

arXiv:2607.08077v1 Announce Type: new Abstract: AI developers face a dual-use dilemma. An AI capability that helps one user cure a disease can help another synthesize one. This dilemma could be resolv

Multimodal Unlearning Across Vision, Language, Video, and Audio: Survey of Methods, Datasets, and Benchmarks

SafetyDGX agent

arXiv:2607.07907v1 Announce Type: cross Abstract: With the growing adoption of VLMs, DMs, LLMs, and AFMs, these multimodal foundation models can inadvertently encode sensitive, copyrighted, biased, or

New research from Meta. (bookmark it) It's on how to fix agents that forget previously made decisions. It's well know that long-horizon agen…

Model ReleasesDGX agent

New research from Meta. (bookmark it) It's on how to fix agents that forget previously made decisions. It's well know that long-horizon agents keep forgetting decisions they already made. Meta researc

@nickfrosst Full podcast > https://www.youtube.com/watch?v=pc4vT8xcSaY

Model ReleasesDGX agent

Nick Frost, likely a notable figure in AI or machine learning, discusses his work and insights in a full-length podcast available on YouTube. The podcast was shared by Cohere, an AI company, suggestin

← Previous
1…626627628629630…1059
Next →