AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlog
87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
Model Releases

PixelPonder: Dynamic Patch Adaptation for Enhanced Multi-Conditional Text-to-Image Generation

DGX agent

arXiv:2503.06684v3 Announce Type: replace Abstract: Recent advances in diffusion-based text-to-image generation have demonstrated promising results through visual condition control. However, existing

model-releasesarxiv-cs-cv
25 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Pope Leo calls for being ‘profoundly human’ in the age of AI

DGX agent

Pope Leo XIV warned of the risks of AI and unconstrained technological power in his first major papal document released on Monday. Magnifica Humanitas is the pope's manifesto on 'safeguarding the huma

model-releasesthe-verge-ai
25 May 2026
Model Releases

PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Negotiations

DGX agent

arXiv:2605.22855v1 Announce Type: cross Abstract: Personalized pricing negotiations are a challenging testbed for LLM agents because successful interaction does not guarantee profitable decision makin

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Push Your Agent: Measuring and Enforcing Quantitative Goal Persistence in Long-Horizon LLM Agents

DGX agent

arXiv:2605.23574v1 Announce Type: new Abstract: Long-horizon language agents can make many plausible local tool calls yet fail to persist until a requested count is actually complete. We study this ga

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

R^3L: Reflect-then-Retry Reinforcement Learning with Language-Guided Exploration, Pivotal Credit, and Positive Amplification

DGX agent

arXiv:2601.03715v2 Announce Type: replace-cross Abstract: Reinforcement learning drives recent advances in LLM reasoning and agentic capabilities, yet current approaches struggle with both exploration

model-releasesarxiv-cs-ai
25 May 2026
Research

RADAR: Relative Angular Divergence Across Representations

DGX agent

arXiv:2605.23028v1 Announce Type: cross Abstract: Machine learning methods rely on data. However, gathering suitable data can be challenging due to availability constraints, cost, or the need for doma

researcharxiv-cs-cl
25 May 2026
Research

Rethinking Transfer Learning for Industrial Inspection: DINOv3 vs. ImageNet Pretraining Across RGB and X-ray Tasks

DGX agent

arXiv:2605.23472v1 Announce Type: new Abstract: Vision foundation models pretrained on web-scale data have recently shown strong transfer capabilities on many downstream tasks, but their effectiveness

researcharxiv-cs-cv
25 May 2026
Model Releases

RMA: an Agentic System for Research-Level Mathematical Problems

DGX agent

arXiv:2605.22875v1 Announce Type: new Abstract: We present extbf{Research Math Agents (RMA)}, an agentic framework for automated reasoning on research-level mathematical problems. Unlike prior studies

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

RoboSurg-VQA: A Multimodal Benchmark for Surgical Segmentation-Aware Visual Question Answering

DGX agent

arXiv:2605.23068v1 Announce Type: new Abstract: Reliable visual understanding in robot-assisted and minimally invasive surgery (RMIS/MIS) demands more than accurate masks: in clinical practice, clinic

model-releasesarxiv-cs-cv
25 May 2026
Research

Robust Counterfactual Inference in Markov Decision Processes

DGX agent

arXiv:2502.13731v5 Announce Type: replace Abstract: This paper addresses a key limitation in existing counterfactual inference methods for Markov Decision Processes (MDPs). Current approaches assume a

researcharxiv-cs-ai
25 May 2026
Model Releases

SciAtlas: A Large-Scale Knowledge Graph for Automated Scientific Research

DGX agent

arXiv:2605.22878v1 Announce Type: new Abstract: The exponential growth of global academic output has confronted researchers and AI agents with an unprecedented ``information explosion,'' where fragmen

model-releasesarxiv-cs-ai
25 May 2026
Research

Smoothed Elicitation Complexity for Approximate Gamma-calibration of Discrete Classification Tasks

DGX agent

arXiv:2605.23017v1 Announce Type: new Abstract: One prominent method of evaluating machine learning model trustworthiness is the notion of calibration. In the binary outcome setting, a probabilistic p

researcharxiv-cs-lg
25 May 2026
Research

SPACENUM: Revisiting Spatial Numerical Understanding in VLMs

DGX agent

arXiv:2605.23898v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly deployed in embodied environments, where they need produce numerical outputs such as action magnitudes an

researcharxiv-cs-ai
25 May 2026
Applications

Sparser Block-Sparse Attention via Token Permutation

DGX agent

arXiv:2510.21270v2 Announce Type: replace-cross Abstract: Scaling the context length of large language models (LLMs) offers significant benefits but is computationally expensive. This expense stems pr

applicationsarxiv-cs-ai
25 May 2026
Safety

SpinFlow: A Physics-Informed Spin Field Framework for Traffic Phase Inference and Transition Detection

DGX agent

arXiv:2605.23306v1 Announce Type: cross Abstract: Active traffic management (ATM) is frequently hindered by traditional macroscopic models and rigid empirical thresholds that fail to capture metastabl

safetyarxiv-cs-lg
25 May 2026
Research

SSDAU: Structured Semantic Data Augmentation for Joint Entity and Relation Extraction

DGX agent

arXiv:2605.23440v1 Announce Type: cross Abstract: Joint Entity and Relation Extraction (JERE) is highly susceptible to weak generalization due to low-quality training data. Data augmentation is a comm

researcharxiv-cs-ai
25 May 2026
Model Releases

StereoGenBench: A Synthetic Multi-Camera Benchmark for Stereo Generation under Controlled Baseline Regimes

DGX agent

arXiv:2605.23237v1 Announce Type: new Abstract: Stereo image and video generation, stereo geometry estimation, and condition-controlled view synthesis require paired data in which the variables that d

model-releasesarxiv-cs-cv
25 May 2026
Research

TCAP: Tri-Component Attention Profiling for Unsupervised Backdoor Detection in MLLM Fine-Tuning

DGX agent

arXiv:2601.21692v2 Announce Type: replace Abstract: Fine-Tuning-as-a-Service (FTaaS) facilitates the customization of Multimodal Large Language Models (MLLMs) but introduces critical backdoor risks vi

researcharxiv-cs-ai
25 May 2026
Applications

The Deterministic Horizon: Impossibility Results as Design Specifications for Trustworthy AI Systems

DGX agent

arXiv:2605.23024v1 Announce Type: new Abstract: Large language models now write software, draft legal documents, and produce clinical notes, yet fundamental limits, from Turing and Arrow to the No Fre

applicationsarxiv-cs-ai
25 May 2026
Model Releases

Vector Retrieval with Similarity and Diversity: How Hard Is It?

DGX agent

arXiv:2407.04573v4 Announce Type: replace-cross Abstract: Dense vector retrieval is an important building block of modern machine learning systems, underlying applications ranging from semantic search

model-releasesarxiv-cs-cl
25 May 2026
Safety

VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction

DGX agent

arXiv:2602.12579v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a dominant paradigm for enhancing Large Language Models (LLMs) reasoning,

safetyarxiv-cs-ai
25 May 2026
Model Releases

What Training Data Teaches RL Memory Agents: An Empirical Study of Curriculum Effects in Memory-Augmented QA

DGX agent

arXiv:2605.23067v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a viable recipe for training LLM agents to reason over external memory banks in multi-session dialogue. Exist

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

When Good Equations Get Bad Scores: Improving Symbolic Regression Through Better Parameter Optimization

DGX agent

arXiv:2605.23272v1 Announce Type: cross Abstract: Symbolic Regression (SR) plays a central role in scientific knowledge discovery by distilling mathematical equations from observational data. Most exi

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Wordle 1,801 4/6 🟨🟨⬛⬛⬛ ⬛🟩⬛⬛⬛ ⬛🟩🟩🟨⬛ 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result where the player solved puzzle #1,801 in four attempts using color-coded feedback (yellow for correct letters in wrong positions, green for correct letters in

model-releasesanthropic--x
25 May 2026
Model Releases

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention

DGX agent

arXiv:2502.04230v3 Announce Type: replace-cross Abstract: The rapid proliferation of generative audio synthesis and editing technologies has raised serious concerns about copyright infringement, data

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

🇺🇸🇪🇺 A new study from the US-based New England Journal of Medicine found that Americans die earlier across all income levels compared to…

DGX agent

🇺🇸🇪🇺 A new study from the US-based New England Journal of Medicine found that Americans die earlier across all income levels compared to their European counterparts. What’s especially notable is that

model-releasesyann-lecun--x
24 May 2026
Model Releases

Aleph 2.0 will blow your mind

DGX agent

Aleph 2.0 will blow your mind Just tested Runway Aleph 2.0 and this blew my mind a bit lol I saw a video like this when Aleph first released and with the new Aleph 2.0 I wanted to create my own versio

model-releasescristobal-valenzuela--x
24 May 2026
Model Releases

An interesting work on Physical AI: PhysX-Omni. First unified sim-ready generation framework for rigid, deformable, and articulated objects,…

DGX agent

An interesting work on Physical AI: PhysX-Omni. First unified sim-ready generation framework for rigid, deformable, and articulated objects, with a diverse dataset and new benchmark. 🌐 https://physx-o

model-releasesclem-delangue--x
24 May 2026
Model Releases

Built an AI screen memory using llama.cpp + Gemma 4 — remembers everything you do on your computer,search/chat or make agents over it. 100% local

DGX agent

This project demonstrates a local AI system built with llama.cpp and Gemma 4 that captures and analyzes screen activity to create persistent memory of user computer interactions, enabling search, chat

model-releasesr-ollama
24 May 2026
Model Releases

It makes many online spaces intolerable. If I want to talk to ChatGPT or Claude, I'll just talk to ChatGPT or Claude, I don't need to talk t…

DGX agent

It makes many online spaces intolerable. If I want to talk to ChatGPT or Claude, I'll just talk to ChatGPT or Claude, I don't need to talk to ChatGPT and Claude pretending to be DoofWarrior123 on X wi

model-releasesethan-mollick--x
24 May 2026
Model Releases

It’s no longer just AI companies & their founders being sued over AI training - individual researchers are now being sued, too. In a new law…

DGX agent

It’s no longer just AI companies & their founders being sued over AI training - individual researchers are now being sued, too. In a new lawsuit, two authors allege that Guillaume Lample, while an AI

model-releasesgary-marcus--x
24 May 2026
Model Releases

Mad House — Usborne Creepy Computer Games

DGX agent

Tool: Mad House — Usborne Creepy Computer Games Via Hacker News I learned that UK publisher Usborne published free PDFs of their 1980s Computer Books, some of which I remember working through on my Co

model-releasessimon-willison
24 May 2026
Model Releases

People often ask what my biggest tip is for getting the most out of Claude Code. These days my #1 tip is: use auto mode Auto mode means no m…

DGX agent

People often ask what my biggest tip is for getting the most out of Claude Code. These days my #1 tip is: use auto mode Auto mode means no more permission prompts. It is the key building block for mul

model-releasesboris-cherny--x
24 May 2026
Model Releases

quick summary of someone's github, cool! here's me

DGX agent

quick summary of someone's github, cool! here's me I always wanted a GitHub dashboard: See my repos, open Issues/PRs, what version I released last, how many commits since last release. So I built one

model-releasesyohei-nakajima--x
24 May 2026
Model Releases

Wordle 1,799 4/6 ⬛⬛⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛🟨⬛⬛⬛ 🟩🟩🟩🟩🟩

DGX agent

This post shows a Wordle game result where the player solved puzzle #1,799 in 4 attempts, with the final answer being a five-letter word where all letters are in the correct positions (indicated by th

model-releasesanthropic--x
24 May 2026
Model Releases

Wordle 1,800 4/6 ⬛⬛⬛⬛🟩 ⬛🟩🟨⬛⬛ 🟩🟩🟨⬛🟩 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result where the player solved puzzle #1,800 in 4 attempts, using the color-coded feedback system (gray for wrong letters, yellow for correct letters in wrong positio

model-releasesanthropic--x
24 May 2026
Model Releases

1. Agreed w @scaling01 that Mythos appears to be better GPT 5.5 on many metrics. 2. Mythos is definitely a major wakeup call wrt security, a…

DGX agent

1. Agreed w @scaling01 that Mythos appears to be better GPT 5.5 on many metrics. 2. Mythos is definitely a major wakeup call wrt security, and will pose problems for real-world systems that aren’t wel

model-releasesgary-marcus--x
23 May 2026
Model Releases

An entropy formula for the Deep Linear Network

DGX agent

arXiv:2509.09088v3 Announce Type: replace Abstract: We study the Riemannian geometry of the Deep Linear Network (DLN) as a foundation for a thermodynamic description of the learning process. The main

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning

DGX agent

arXiv:2605.21968v1 Announce Type: new Abstract: Optimization is essential in deep learning. The foundational method upon which most optimizers are built is momentum-based stochastic gradient descent.

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

ASSEMBLAGE-DEEPHISTORY: A Cross-Build Binary Dataset with Temporal Coverage

DGX agent

arXiv:2605.21615v1 Announce Type: cross Abstract: Existing binary corpora typically capture only one or two axes of binary variation: they either provide cross-compiler builds without a temporal axis,

model-releasesarxiv-cs-lg
23 May 2026
Local Ai

b9296

DGX agent

b9296 is a build release of llama.cpp , the C/C++ implementation for efficient large language model inference. As an intermediate build in the llama.cpp release cycle, it likely includes recent bug fi

local-aillama-cpp-releases
23 May 2026
Model Releases

Chebyshev Policies and the Mountain Car Problem: Reinforcement Learning for Low-Dimensional Control Tasks

DGX agent

arXiv:2605.22305v1 Announce Type: new Abstract: We analytically solve the Mountain Car problem, a canonical benchmark in RL, and derive an optimal control solution, closing a gap after 36 years. This

model-releasesarxiv-cs-lg
23 May 2026
Tools

co-sign. a very handy mental framework for what kinds of learning transformers do well today, and why it runs into limitations. when @ankit2…

DGX agent

co-sign. a very handy mental framework for what kinds of learning transformers do well today, and why it runs into limitations. when @ankit2119 and i wrote about the need for adversarial world models

toolsswyx--x
23 May 2026
Model Releases

Cross-domain benchmarks reveal when coordinated AI agents improve scientific inference from partial evidence

DGX agent

arXiv:2605.22300v1 Announce Type: cross Abstract: Scientific evidence often spans instruments, databases, and disciplines, so no single source records the full phenomenon. This makes it difficult to d

model-releasesarxiv-cs-lg
23 May 2026
Safety

Cross-Species RSA Reveals Conserved Early Visual Alignment but Divergent Higher-Area Rankings Across Human fMRI and Macaque Electrophysiology

DGX agent

arXiv:2605.22401v1 Announce Type: new Abstract: Does the relationship between learning rules and brain alignment generalize across species? We extend our prior finding that untrained CNNs match backpr

safetyarxiv-cs-lg
23 May 2026
Model Releases

Do Not Trust The Auctioneer: Learning to Bid in Feedback-Manipulated Auctions

DGX agent

arXiv:2605.22438v1 Announce Type: cross Abstract: Shilling is the use of artificial bids to make competition appear stronger and push prices upward. We study repeated first-price auctions in which shi

model-releasesarxiv-cs-lg
23 May 2026
Research

Double descent for least-squares interpolation on contaminated data: A simulation study

DGX agent

arXiv:2605.21494v1 Announce Type: new Abstract: Overparametrized models can exhibit an excellent generalization performance, although they should be prone to overfitting according to classical statist

researcharxiv-cs-lg
23 May 2026
Model Releases

Dropout Universality: Scaling Laws and Optimal Scheduling at the Edge-of-Chaos

DGX agent

arXiv:2605.21648v1 Announce Type: new Abstract: We develop a mean-field theory of dropout as a perturbation of critical signal propagation at the edge of chaos. Dropout shifts the perfect-alignment fi

model-releasesarxiv-cs-lg
23 May 2026
← Previous
1…844845846847848…1314
Next →