AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,612 results
Model Releases

Steve Bannon and 60+ Trump allies sign a Humans First-led letter urging Trump to mandate government testing and approval of powerful AI models before release (Ashley Gold/Axios)

DGX agent

Ashley Gold / Axios: Steve Bannon and 60+ Trump allies sign a Humans First-led letter urging Trump to mandate government testing and approval of powerful AI models before release — A group of more tha

model-releasestechmeme
18 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

StippleDiffusion: Capacity-Constrained Stippling using Controlled Diffusion

DGX agent

arXiv:2605.15816v1 Announce Type: cross Abstract: Stipple patterns, point sets whose local density tracks a target image, are traditionally produced by per-density iterative optimizers, which are slow

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model

DGX agent

arXiv:2605.15733v1 Announce Type: cross Abstract: Humans abstract experiences into structured representations to facilitate pattern inference and knowledge transfer. While the hippocampal-entorhinal (

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Structure-BiEval: A Self-Supervised, Dual-Track Framework for Decoupling Structure and Content in LLM Evaluation for Web Information Systems

DGX agent

arXiv:2601.19923v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into the core of Web-based autonomous agents and complex Web Information Systems, their ability to fait

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

STS: Efficient Sparse Attention with Speculative Token Sparsity

DGX agent

arXiv:2605.15508v1 Announce Type: cross Abstract: The quadratic complexity of attention imposes severe memory and computational bottlenecks on Large Language Model (LLM) inference. This challenge is p

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

SurvivalPFN: Amortizing Survival Prediction via In-Context Bayesian Inference

DGX agent

arXiv:2605.15488v1 Announce Type: new Abstract: Survival analysis provides a powerful statistical framework for modeling time-to-event outcomes in the presence of censoring. However, selecting an appr

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception

DGX agent

arXiv:2602.21141v2 Announce Type: replace Abstract: Object perception is fundamental for tasks such as robotic material handling and quality inspection. However, modern supervised deep-learning models

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

T2T-LA: A Topology-to-Topology LLM Agent for Graph Learning with Neither Feature Access nor Task Knowledge

DGX agent

arXiv:2512.08964v4 Announce Type: replace Abstract: Graph learning aims to convert data into graph representations, which are fundamental to many problems in machine learning for CAD, where circuits,

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

TACO: General Acrobatic Flight Control via Target-and-Command-Oriented Reinforcement Learning

DGX agent

arXiv:2503.01125v4 Announce Type: replace Abstract: Although acrobatic flight control has been studied extensively, one key limitation of the existing methods is that they are usually restricted to sp

model-releasesarxiv-cs-ro
18 May 2026
Model Releases

Tadpole: Autoencoders as Foundation Models for 3D PDEs with Online Learning

DGX agent

arXiv:2605.15284v1 Announce Type: new Abstract: We introduce Tadpole, a novel foundation model for three-dimensional partial differential equations (PDEs) that addresses key challenges in transferabil

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

There are a lot of coding and reasoning benchmarks for AI agents, but not a lot for document understanding - which is a prerequisite for all…

DGX agent

There are a lot of coding and reasoning benchmarks for AI agents, but not a lot for document understanding - which is a prerequisite for all downstream knowledge work. We released ParseBench ~a month

model-releasesjerry-liu--x
18 May 2026
Model Releases

These kids are serial criminals with a callous disregard for life. If they are ever released from jail they will surely harm again. Austin P…

DGX agent

These kids are serial criminals with a callous disregard for life. If they are ever released from jail they will surely harm again. Austin PD, Travis Co. Sheriff Office & Manor PD did their job. Texas

model-releaseselon-musk--x
18 May 2026
Model Releases

💯 this is why I really like Learning mode in Claude Code I personally use this for all my side projects and it keeps me so much sharper, gr…

DGX agent

💯 this is why I really like Learning mode in Claude Code I personally use this for all my side projects and it keeps me so much sharper, great if you want to use Claude Code but still stay hands-on! /

model-releasesboris-cherny--x
18 May 2026
Model Releases

Today in AI Engineering (May 17) • Nous Research ships Hermes Agent v0.14.0: Grok subs, Codex runtime, Windows beta • LangSmith Engine relea…

DGX agent

Today in AI Engineering (May 17) • Nous Research ships Hermes Agent v0.14.0: Grok subs, Codex runtime, Windows beta • LangSmith Engine releases trace issue clustering, drafts PRs and evals from prod t

model-releasesharrison-chase--x
18 May 2026
Model Releases

TokenButler: Token Importance is Predictable

DGX agent

arXiv:2503.07518v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) rely on the Key-Value (KV) Cache to store token history, enabling efficient decoding of tokens. As the KV-Cache g

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression

DGX agent

arXiv:2602.08324v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning successfully enhances the reasoning capabilities of Large Language Models (LLMs), yet it incurs substantial computa

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Transformer Scalability Crisis: The First Comprehensive Empirical Analysis of Performance Walls in Modern Language Models

DGX agent

arXiv:2605.15413v1 Announce Type: new Abstract: Despite the remarkable success of transformer architectures in natural language processing, their scalability limitations remain poorly understood throu

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

TSBOW -- Traffic Surveillance Benchmark for Occluded Vehicles Under Various Weather Conditions

DGX agent

arXiv:2602.05414v2 Announce Type: replace Abstract: Global warming has intensified the frequency and severity of extreme weather events, which degrade CCTV signal and video quality while disrupting tr

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Tube Loss: A Novel Approach for Prediction Interval Estimation

DGX agent

arXiv:2412.06853v4 Announce Type: replace-cross Abstract: This paper proposes a novel loss function, called 'Tube Loss', for simultaneous estimation of bounds of a Prediction Interval (PI) in the regr

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

UAM: A Dual-Stream Perspective on Forgetting in VLA Training

DGX agent

arXiv:2605.15735v1 Announce Type: cross Abstract: Vision--language--action (VLA) models are typically built by fine-tuning a pretrained vision--language model (VLM) on action data. However, we show th

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Unlocking Dense Metric Depth Estimation in VLMs

DGX agent

arXiv:2605.15876v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at 2D tasks such as grounding and captioning, yet remain limited in 3D understanding. A key limitation is their text

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

DGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

VideoGameBench: Can Vision-Language Models complete popular video games?

DGX agent

arXiv:2505.18134v3 Announce Type: replace Abstract: Vision-language models (VLMs) have achieved strong results on coding and math benchmarks that are challenging for humans, yet their ability to perfo

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation

DGX agent

arXiv:2605.16079v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have shown significant progress in video understanding, yet they face substantial challenges in tasks requiring p

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

VideoVerse: Does Your T2V Generator Have World Model Capability to Synthesize Videos?

DGX agent

arXiv:2510.08398v4 Announce Type: replace Abstract: The recent rapid advancement of Text-to-Video (T2V) generation technologies are engaging the trained models with more world model ability, making th

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

We causally trained a lot of SOTA search models internally, shall we make some small release from time to time 🤣🤣

DGX agent

We causally trained a lot of SOTA search models internally, shall we make some small release from time to time 🤣🤣 @bo_wangbo stealth releasing probably the strongest open multilingual ColBERT (and it'

model-releasesclem-delangue--x
18 May 2026
Model Releases

'We give you model choice, without infrastructure chaos' — @MichaelDell, live from #DellTechWorld 🎤 Kimi K2.6, DeepSeek V4 Pro, GLM 5.1, Mi…

DGX agent

'We give you model choice, without infrastructure chaos' — @MichaelDell, live from #DellTechWorld 🎤 Kimi K2.6, DeepSeek V4 Pro, GLM 5.1, MiniMax M2.7 & DeepSeek V4 Flash are now one click away on Dell

model-releasesclem-delangue--x
18 May 2026
Model Releases

WeatherOcc3D: VLM-Assisted Adverse Weather Aware 3D Semantic Occupancy Prediction

DGX agent

arXiv:2605.16127v1 Announce Type: new Abstract: While multi-modal 3D semantic occupancy prediction typically enhances robustness by fusing camera and LiDAR inputs, its effectiveness is fundamentally c

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Weight Concentration Regularization for Improving Pruning Robustness Under High Sparsity

DGX agent

arXiv:2511.14282v2 Announce Type: replace-cross Abstract: Deep neural networks achieve outstanding performance across vision and language tasks, yet their large parameter counts limit deployment in re

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Well, that's a Turing Test of a sort. (But gosh is the AI writing obvious if you use these systems at all - and this is obviously ChatGPT wr…

DGX agent

Well, that's a Turing Test of a sort. (But gosh is the AI writing obvious if you use these systems at all - and this is obviously ChatGPT writing, not Claude) Well, this is a first: a ChatGPT-generate

model-releasesethan-mollick--x
18 May 2026
Model Releases

What are best practices for running Claude Code at scale? New blog post on what we've learned from teams running it across multi-million-lin…

DGX agent

What are best practices for running Claude Code at scale? New blog post on what we've learned from teams running it across multi-million-line monorepos, decades-old legacy systems, and distributed mic

model-releasesboris-cherny--x
18 May 2026
Model Releases

What we announced in streaming AI at Next ‘26

DGX agent

Every device, user, and microservice generates data. Ingesting this data, extracting meaning and insights, and driving business decisions in real time has the potential to deliver transformational bus

model-releasesgoogle-cloud-ai
18 May 2026
Model Releases

xAI has Released a blog on Skills

DGX agent

xAI has released a blog post discussing skills, likely covering topics related to AI capabilities, competencies, or skill development within their AI systems. The announcement was shared by Elon Musk

model-releaseselon-musk--x
18 May 2026
Model Releases

XSearch: Explainable Code Search via Concept-to-Code Alignment

DGX agent

arXiv:2605.16046v1 Announce Type: cross Abstract: Semantic code search has been widely adopted in both academia and industry. These approaches embed natural-language queries and code snippets into a s

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

You can now create more with Claude Design. We've doubled token limits across every plan.

DGX agent

Anthropic has doubled token limits across all Claude plans, enabling users to work with larger amounts of text and create more complex projects. This upgrade applies to the Claude Design tool and repr

model-releasesthariq--x
18 May 2026
Model Releases

Zero-Shot Goal Recognition with Large Language Models

DGX agent

arXiv:2605.15333v1 Announce Type: new Abstract: Large language models have recently reached near-parity with classical planners on well-known planning domains, yet this competence relies on world-know

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Cl…

DGX agent

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Claude in 5 min 2026 → train your own model over a weekend hap

model-releasesfireworks-ai--x
17 May 2026
Model Releases

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent …

DGX agent

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent models, let alone recent agentic tools. 'The Cybernetic Team

model-releasesethan-mollick--x
17 May 2026
Model Releases

DeepSeek Exposed: Users Can Access Each Other's Conversations with a Special Input[D]

DGX agent

A vulnerability in DeepSeek's website exposed a significant amount of data, including user chats. A publicly accessible ClickHouse database belonging to DeepSeek allowed full control over database ope

model-releasesr-machinelearning
17 May 2026
Model Releases

doomers and AI fans love hearing “it’s over” - but @scaling01 really should have noted that the below refers to a domain specific benchmark …

DGX agent

doomers and AI fans love hearing “it’s over” - but @scaling01 really should have noted that the below refers to a domain specific benchmark (exploitbench) rather than general purpose AI. also note the

model-releasesgary-marcus--x
17 May 2026
Model Releases

G4-MeroMero-31B-uncensored-heretic is Out Now, A finetune of Gemma 4 31B it designed for creative tasks, with KLD of 0.0100 and 15/100 Refusals!

DGX agent

G4-MeroMero-31B-uncensored-heretic is a fine-tuned variant of Gemma 4 31B optimized for creative tasks, featuring low KL divergence (0.0100) and minimal refusals (15/100). The model is designed to be

model-releasesr-ollama
17 May 2026
Model Releases

Gemini for Science: AI experiments and tools for a new era of discovery

DGX agent

Gemini for Science is a collection of science tools and experiments designed to expand the scale and precision of scientific exploration , including tools to help scientists explore more hypotheses, v

model-releasesgoogle-deepmind
17 May 2026
Model Releases

GPT-5.5 Pro faces its hardest academic challenge: to apply the technique from a paper analyzing which word pairs were funny & why to come up…

DGX agent

GPT-5.5 Pro faces its hardest academic challenge: to apply the technique from a paper analyzing which word pairs were funny & why to come up with its own It came up with scrotum snorkel, tuba subpoena

model-releasesethan-mollick--x
17 May 2026
Model Releases

Harness profiles! https://docs.langchain.com/oss/python/deepagents/profiles

DGX agent

Harness profiles! https://docs.langchain.com/oss/python/deepagents/profiles I like what Langchain recently released in their deepagents harness, which is an adapter to modify the syntax of primitive f

model-releasesharrison-chase--x
17 May 2026
Model Releases

Introducing Gemini Omni

DGX agent

Gemini Omni is a multimodal AI model that allows users to create content from any input and edit naturally using conversational language . Users can combine images, audio, video, and text as input to

model-releasesgoogle-deepmind
17 May 2026
Model Releases

Making it easier to understand how content was created and edited

DGX agent

Google DeepMind expanded content transparency and verification tools to help users understand how content was created and edited across the web, integrating these capabilities into Search, Gemini, Chr

model-releasesgoogle-deepmind
17 May 2026
Model Releases

There’s an open question on whether grep is all you need for agentic search. This recent paper by @PwCUS (Sen et al.) seems to suggest that.…

DGX agent

There’s an open question on whether grep is all you need for agentic search. This recent paper by @PwCUS (Sen et al.) seems to suggest that. It’s titled “Is Grep All You Need? How Agent Harnesses Resh

model-releasesjerry-liu--x
17 May 2026
Model Releases

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the env…

DGX agent

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the environment itself is the verifier. 🏆Among frontier models, @An

model-releaseszhipu-ai--x
17 May 2026
← Previous
1…306307308309310…472
Next →