AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,767 results
Model Releases

Grep timeout issue fixed in latest Grok Build

DGX agent

Grep timeout issue fixed in latest Grok Build Grok Build update just released v0.2.31 Release Notes: Bug Fixes: • Marketplace skills without proper descriptions are now hidden from listings instead of

model-releaseselon-musk--x
7 Jun 2026
Model Releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Have been extensively testing Claude Workflows this weekend, with the best model possible. Threw it at my whole code base, combing for bugs.…

DGX agent

Have been extensively testing Claude Workflows this weekend, with the best model possible. Threw it at my whole code base, combing for bugs. 144 found and fixed! Geez... It is a large code base, for s

model-releasesboris-cherny--x
7 Jun 2026
Model Releases

LOL. Even Claude sees through Hinton’s nonsense. (though see the articles I posted earlier, for converging sources I put more weight on)

DGX agent

LOL. Even Claude sees through Hinton’s nonsense. (though see the articles I posted earlier, for converging sources I put more weight on) This is my prompt and this is Claude's response: ME: What do yo

model-releasesgary-marcus--x
7 Jun 2026
Model Releases

see eg https://www.scmp.com/tech/tech-trends/article/3271858/ai-race-alibaba-tencent-quickly-adopt-metas-new-llama-31-model-amid-excitement …

DGX agent

see eg https://www.scmp.com/tech/tech-trends/article/3271858/ai-race-alibaba-tencent-quickly-adopt-metas-new-llama-31-model-amid-excitement and https://medium.com/the-endless-forge/zuckerbergs-llama-f

model-releasesgary-marcus--x
7 Jun 2026
Model Releases

The first wave of AI-native applications is wrapping tokens and providing in-app agents. As agent usage centralizes around core apps (e.g. C…

DGX agent

The first wave of AI-native applications is wrapping tokens and providing in-app agents. As agent usage centralizes around core apps (e.g. Claude Code, Codex), there's this emerging wave of building s

model-releasesjerry-liu--x
7 Jun 2026
Model Releases

There was an inflection point recently where the tide shifted to model pickers and OSS Mix of tokenmaxxing/cost fatigue, nemotron coalition,…

DGX agent

There was an inflection point recently where the tide shifted to model pickers and OSS Mix of tokenmaxxing/cost fatigue, nemotron coalition, harness step functions, brains/claws, etc Long live those w

model-releasesclem-delangue--x
7 Jun 2026
Model Releases

Wasn't Krea 2 supposed to be released ?

DGX agent

Krea 2, Krea's first foundation image model built from scratch, was announced on May 12, 2026 , with a focus on aesthetics, style transfer, and creative control . Krea 2 became available to everyone s

model-releasesr-stablediffusion
7 Jun 2026
Model Releases

Wordle 1,813 4/6 ⬛⬛⬛🟨⬛ ⬛⬛⬛⬛🟨 🟨🟨🟨⬛⬛ 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result shared by Anthropic on X (Twitter), showing puzzle #1,813 solved in 4 attempts with a final correct answer displayed through colored emoji tiles (green indicat

model-releasesanthropic--x
7 Jun 2026
Model Releases

Zuckerberg and LeCun’s unilateral decision to open source Llama likely (partly) catalyzed China’s AI industry — and may have done truly mass…

DGX agent

Zuckerberg and LeCun’s unilateral decision to open source Llama likely (partly) catalyzed China’s AI industry — and may have done truly massive harm to American business interests. We are now starting

model-releasesgary-marcus--x
7 Jun 2026
Model Releases

ADK Arena: Evaluating Agent Development Kits via LLM-as-a-Developer

DGX agent

arXiv:2606.05548v1 Announce Type: cross Abstract: The rapid proliferation of Agent Development Kits (ADKs), SDK-level frameworks for building LLM-powered autonomous agents, has outpaced any empirical

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Agent Memory: Characterization and System Implications of Stateful Long-Horizon Workloads

DGX agent

arXiv:2606.06448v1 Announce Type: new Abstract: LLM agents are increasingly deployed on long-horizon tasks requiring sustained reasoning over extended interaction histories. Realizing this at scale re

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Agent-Orchestrated Adaptive RAG: A Comparative Study on Structured and Multi-Hop Retrieval

DGX agent

arXiv:2606.05658v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding their responses in external knowledge, but conventional pipeli

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Agentic Monte Carlo: Simulating Reinforcement Learning for Black-Box Agents

DGX agent

arXiv:2606.05296v1 Announce Type: cross Abstract: LLM agents operate in two distinct regimes: open-weight agents amenable to reinforcement learning (RL) and black-box agents whose behaviour must be co

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

AttackPathGNN: Cross-function vulnerability detection in smart contracts using state interference graphs and conjunction pooling

DGX agent

arXiv:2606.05986v1 Announce Type: cross Abstract: Existing learning-based detectors for Solidity smart-contracts reduce vulnerability detection to syntactic pattern matching within single functions, y

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Benchmark Everything Everywhere All at Once

DGX agent

arXiv:2606.06462v1 Announce Type: new Abstract: Benchmarks are fundamental for evaluating and advancing LLMs and MLLMs by providing standardized and explicit measures of performance. However, their co

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Benchmarking Counterfactual Prediction in Epidemic Time Series with Time-Varying Interventions

DGX agent

arXiv:2606.05692v1 Announce Type: cross Abstract: Deep learning has enabled significant advances in time-series causal inference, yet progress remains constrained by the lack of realistic benchmarks w

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Beyond Output Matching: Preserving Internal Geometry in NVFP4 LLM Distillatio

DGX agent

arXiv:2606.05682v1 Announce Type: new Abstract: Demand for low-precision inference, including NVFP4-based approaches, has grown as large language models are increasingly deployed in latency and cost c

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

🚨BREAKING: Miguel Bosé, the biggest Spanish-language pop star of the last few decades, has just released a video taking a knee and putting …

DGX agent

🚨BREAKING: Miguel Bosé, the biggest Spanish-language pop star of the last few decades, has just released a video taking a knee and putting his hand over his heart in honour of Henry Nowak This has now

model-releaseselon-musk--x
6 Jun 2026
Model Releases

Brick-Composer: Using MLLMs for Assembly with Diverse Bricks

DGX agent

arXiv:2606.05445v1 Announce Type: new Abstract: We dream of AI agents that can read arbitrary designs and construct real-world objects from reusable building blocks. As a first step toward this vision

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Can AI Refute Economic Theory? Evidence from Beyond the Knowledge Cutoff

DGX agent

arXiv:2606.05383v1 Announce Type: cross Abstract: Can artificial intelligence (AI) refute economic theory? I document experiments in which I asked several AI models (Gemini, Refine, Claude, and ChatGP

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Can LLMs Write Correct TLA+ Specifications? Evaluating Natural-Language-to-TLA+ Generation

DGX agent

arXiv:2606.05792v1 Announce Type: new Abstract: TLA+ has supported industrial verification at companies such as Amazon and Microsoft, yet writing correct TLA+ specifications from natural language stil

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

CangLing-KnowFlow: A Unified Knowledge-and-Flow-fused Agent for Comprehensive Remote Sensing Applications

DGX agent

arXiv:2512.15231v3 Announce Type: replace Abstract: The automated and intelligent processing of massive remote sensing (RS) datasets is critical in Earth observation (EO). Existing automated systems a

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs

DGX agent

arXiv:2606.05966v1 Announce Type: cross Abstract: Understanding and reasoning about the physical world is the foundation of intelligent behavior, yet state-of-the-art vision-language models (VLMs) sti

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Chip Export rules: made China/Huawei stronger, accomplished relatively little. USG stakes in US AI: will make Mistral and other “sovereign A…

DGX agent

Chip Export rules: made China/Huawei stronger, accomplished relatively little. USG stakes in US AI: will make Mistral and other “sovereign AI” efforts stronger, freak out the rest of the world, escala

model-releasesgary-marcus--x
6 Jun 2026
Model Releases

Closing the Loop on Latent Reasoning via Test-Time Reconstruction

DGX agent

arXiv:2606.06252v1 Announce Type: new Abstract: Recent work moves intermediate reasoning from natural-language traces into latent or cache-level representations to reduce token overhead and avoid a di

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

CogManip: Benchmarking Manipulative Behavior in Multi-Turn Interactions with Large Language Model

DGX agent

arXiv:2606.06099v1 Announce Type: new Abstract: Whether Large Language Models (LLMs) exhibit covert psychological manipulation in complex human-AI interactions has garnered increasing safety concerns.

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Critic-Guided Heterogeneous Multi-Agent Reasoning for Reliable Mathematical Problem Solving

DGX agent

arXiv:2606.05704v1 Announce Type: new Abstract: Recent Large Language Models (LLMs) have shown impressive reasoning abilities; but they are still susceptible to hallucinations, intermediate reasoning

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Cross-Epoch Adaptive Rollout Optimization for RL Post-Training

DGX agent

arXiv:2606.05606v1 Announce Type: cross Abstract: LLM post-training often relies on reinforcement learning methods that sample multiple rollouts per prompt, yet most existing approaches use a fixed ro

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

CTIConnect: A Benchmark for Retrieval-Augmented LLMs over Heterogeneous Cyber Threat Intelligence

DGX agent

arXiv:2510.11974v2 Announce Type: replace-cross Abstract: Cyber Threat Intelligence (CTI) is foundational to modern cybersecurity, enabling organizations to proactively defend against evolving threats

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Data Flow Control: Data Safety Policies for AI Agents

DGX agent

arXiv:2606.05679v1 Announce Type: cross Abstract: Agents increasingly generate SQL, orchestrate pipelines, and automate data analysis on behalf of users. While recent work improves query correctness,

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Do More Agents Help? Controlled and Protocol-Aligned Evaluation of LLM Agent Workflows

DGX agent

arXiv:2606.05670v1 Announce Type: new Abstract: Does adding more agents help an LLM workflow once compared systems share the same benchmark loader, tool access, answer contract, usage accounting, and

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

DPBench: Structural Determinants of Multi-Agent LLM Coordination Under Simultaneous Resource Contention

DGX agent

arXiv:2602.13255v2 Announce Type: replace Abstract: We present DPBench, a benchmark for evaluating coordination in multi-agent systems built from large language models. Existing benchmarks measure tas

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

DragOn: A Benchmark and Dataset for Drag-Based GUI Interactions

DGX agent

arXiv:2606.06322v1 Announce Type: new Abstract: GUI agents - vision-based models that control desktops, web browsers, and mobile devices through graphical user interfaces - promise to automate a wide

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Edit-R2: Context-Aware Reinforcement Learning for Multi-Turn Image Editing

DGX agent

arXiv:2606.05950v1 Announce Type: new Abstract: Text-guided image editing has advanced rapidly with diffusion models and unified multimodal foundation models. However, most existing methods remain con

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Enhancing Software Engineering Through Closed-Loop Memory Optimization

DGX agent

arXiv:2606.05646v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled powerful software engineering (SE) agents capable of navigating complex codebases and resolving real-world i

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Evaluating Agentic Configuration Repair for Computer Networks

DGX agent

arXiv:2606.06212v1 Announce Type: new Abstract: Misconfigurations in computer networks remain a major source of critical Internet outages. Research is turning to Large Language Models (LLMs) to automa

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Evaluation of LLMs for Mathematical Formalization in Lean

DGX agent

arXiv:2606.05632v1 Announce Type: new Abstract: Within the past few years, the ability of Large Language Models (LLMs) to generate formal mathematical proofs has improved drastically. We provide a com

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Exploring LLMs for South Asian Music Understanding and Generation

DGX agent

arXiv:2606.05522v1 Announce Type: cross Abstract: Recent advancements in Large Language Models (LLMs) have shown promising results in music understanding and generation tasks. However, existing works

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Fireworks Training Platform keeps expanding. Leading US open weight model Nemotron 3 Ultra is now ready for post-training: SFT and DPO via L…

DGX agent

Fireworks Training Platform keeps expanding. Leading US open weight model Nemotron 3 Ultra is now ready for post-training: SFT and DPO via LoRA or full-parameter, on the same infrastructure that serve

model-releasesfireworks-ai--x
6 Jun 2026
Model Releases

Four papers out recently: 1. http://political-manipulation.ai: Measures and reduces political bias in LLMs; Claude is especially biased 2. h…

DGX agent

Four papers out recently: 1. http://political-manipulation.ai: Measures and reduces political bias in LLMs; Claude is especially biased 2. http://aibetrayal.com: The public can insert backdoors into A

model-releasesdan-hendrycks--x
6 Jun 2026
Model Releases

GenTI: Benchmarking LLMs for Autonomous IDPS Rule Generation for Unseen Attacks

DGX agent

arXiv:2606.05844v1 Announce Type: cross Abstract: Rule-based Intrusion Detection and Prevention Systems (IDPS) offer precise attack detection as well as mitigation, however their manually crafted, sig

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Geographic Bias and Diversity in AI Evaluation

DGX agent

arXiv:2606.05187v1 Announce Type: cross Abstract: Among the many challenges hindering the responsible development and deployment of AI, arguably none has faced more intense scrutiny than bias in its v

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

GITCO: Gated Inference-Time Context Optimization in TSFMs

DGX agent

arXiv:2606.05332v1 Announce Type: new Abstract: Patch-based Time Series Foundation Models (TSFMs) suffer from context poisoning: structurally anomalous patches capture disproportionate attention and s

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Goedel-Architect: Streamlining Formal Theorem Proving with Blueprint Generation and Refinement

DGX agent

arXiv:2606.06468v1 Announce Type: new Abstract: We introduce Goedel-Architect, an agentic framework for formal theorem proving in Lean 4 centered on blueprint generation and refinement. A blueprint is

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

GuardNet: Ensemble Strategies of Shallow Neural Networks for Robust Prompt Injection and Jailbreak Detection

DGX agent

arXiv:2606.05566v1 Announce Type: new Abstract: Large Language Models (LLMs) have transformed natural language processing, but they remain vulnerable to Prompt Injection (PI) and Jailbreak (JB) attack

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

How Far Did They Go? The Persuasive Tactics of Covert LLM Agents in a Discontinued Field Experiment

DGX agent

arXiv:2606.05256v1 Announce Type: new Abstract: This study analyzes a publicly released dataset from a discontinued field experiment on Reddit's r/ChangeMyView. The intervention, conducted by unknown,

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

I Know What You Meme, Even If it Emerged Today: Understanding Evolving Memes through Open-World Knowledge Acquisition

DGX agent

arXiv:2606.05316v1 Announce Type: new Abstract: Multimodal memes are dynamic and often require up to date background knowledge for interpretation. Existing methods often overlook such knowledge or rel

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

I want AI to succeed, and have been positive about many systems (Claude Code, AlphaFold, AlphaGeometry, Cicero, etc). It’s not AI that I hat…

DGX agent

I want AI to succeed, and have been positive about many systems (Claude Code, AlphaFold, AlphaGeometry, Cicero, etc). It’s not AI that I hate; it’s bullshit and greed that I despise. Not my fault that

model-releasesgary-marcus--x
6 Jun 2026
← Previous
1…209210211212213…475
Next →