AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
25 May 2026

Ontological Knowledge Blocks: Executable Compliance and Profile-Based Validation for Trustworthy AI Systems

Model ReleasesDGX agent

arXiv:2605.23297v1 Announce Type: new Abstract: AI-enabled services deployed in critical digital infrastructure are subject to governance obligations spanning transparency, accountability, fairness, a

Optimization of randomized neural networks for transfer operator approximation

Model ReleasesDGX agent

arXiv:2605.23689v1 Announce Type: new Abstract: RaNNDy is a randomized neural network architecture for the data-driven approximation of transfer operators associated with complex dynamical systems. Th

Order-Optimal Sequential 1-Bit Mean Estimation in General Tail Regimes

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.07796v2 Announce Type: replace-cross Abstract: In this paper, we study the problem of mean estimation under 1-bit communication constraints. We propose a novel adaptive mean estimator based

PathCal: State-Aware Reflection-Marker Calibration for Efficient Reasoning

ResearchDGX agent

arXiv:2605.23074v1 Announce Type: new Abstract: The emergence of Large Reasoning Language Models (LRMs) has paved the way for tackling complex reasoning tasks through test-time scaling by generating l

PixelPonder: Dynamic Patch Adaptation for Enhanced Multi-Conditional Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2503.06684v3 Announce Type: replace Abstract: Recent advances in diffusion-based text-to-image generation have demonstrated promising results through visual condition control. However, existing

Pope Leo calls for being ‘profoundly human’ in the age of AI

Model ReleasesDGX agent

Pope Leo XIV warned of the risks of AI and unconstrained technological power in his first major papal document released on Monday. Magnifica Humanitas is the pope's manifesto on 'safeguarding the huma

PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Negotiations

Model ReleasesDGX agent

arXiv:2605.22855v1 Announce Type: cross Abstract: Personalized pricing negotiations are a challenging testbed for LLM agents because successful interaction does not guarantee profitable decision makin

Push Your Agent: Measuring and Enforcing Quantitative Goal Persistence in Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2605.23574v1 Announce Type: new Abstract: Long-horizon language agents can make many plausible local tool calls yet fail to persist until a requested count is actually complete. We study this ga

R^3L: Reflect-then-Retry Reinforcement Learning with Language-Guided Exploration, Pivotal Credit, and Positive Amplification

Model ReleasesDGX agent

arXiv:2601.03715v2 Announce Type: replace-cross Abstract: Reinforcement learning drives recent advances in LLM reasoning and agentic capabilities, yet current approaches struggle with both exploration

RADAR: Relative Angular Divergence Across Representations

ResearchDGX agent

arXiv:2605.23028v1 Announce Type: cross Abstract: Machine learning methods rely on data. However, gathering suitable data can be challenging due to availability constraints, cost, or the need for doma

Rethinking Transfer Learning for Industrial Inspection: DINOv3 vs. ImageNet Pretraining Across RGB and X-ray Tasks

ResearchDGX agent

arXiv:2605.23472v1 Announce Type: new Abstract: Vision foundation models pretrained on web-scale data have recently shown strong transfer capabilities on many downstream tasks, but their effectiveness

RMA: an Agentic System for Research-Level Mathematical Problems

Model ReleasesDGX agent

arXiv:2605.22875v1 Announce Type: new Abstract: We present extbf{Research Math Agents (RMA)}, an agentic framework for automated reasoning on research-level mathematical problems. Unlike prior studies

RoboSurg-VQA: A Multimodal Benchmark for Surgical Segmentation-Aware Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.23068v1 Announce Type: new Abstract: Reliable visual understanding in robot-assisted and minimally invasive surgery (RMIS/MIS) demands more than accurate masks: in clinical practice, clinic

Robust Counterfactual Inference in Markov Decision Processes

ResearchDGX agent

arXiv:2502.13731v5 Announce Type: replace Abstract: This paper addresses a key limitation in existing counterfactual inference methods for Markov Decision Processes (MDPs). Current approaches assume a

SciAtlas: A Large-Scale Knowledge Graph for Automated Scientific Research

Model ReleasesDGX agent

arXiv:2605.22878v1 Announce Type: new Abstract: The exponential growth of global academic output has confronted researchers and AI agents with an unprecedented ``information explosion,'' where fragmen

Smoothed Elicitation Complexity for Approximate Gamma-calibration of Discrete Classification Tasks

ResearchDGX agent

arXiv:2605.23017v1 Announce Type: new Abstract: One prominent method of evaluating machine learning model trustworthiness is the notion of calibration. In the binary outcome setting, a probabilistic p

SPACENUM: Revisiting Spatial Numerical Understanding in VLMs

ResearchDGX agent

arXiv:2605.23898v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly deployed in embodied environments, where they need produce numerical outputs such as action magnitudes an

Sparser Block-Sparse Attention via Token Permutation

ApplicationsDGX agent

arXiv:2510.21270v2 Announce Type: replace-cross Abstract: Scaling the context length of large language models (LLMs) offers significant benefits but is computationally expensive. This expense stems pr

SpinFlow: A Physics-Informed Spin Field Framework for Traffic Phase Inference and Transition Detection

SafetyDGX agent

arXiv:2605.23306v1 Announce Type: cross Abstract: Active traffic management (ATM) is frequently hindered by traditional macroscopic models and rigid empirical thresholds that fail to capture metastabl

SSDAU: Structured Semantic Data Augmentation for Joint Entity and Relation Extraction

ResearchDGX agent

arXiv:2605.23440v1 Announce Type: cross Abstract: Joint Entity and Relation Extraction (JERE) is highly susceptible to weak generalization due to low-quality training data. Data augmentation is a comm

StereoGenBench: A Synthetic Multi-Camera Benchmark for Stereo Generation under Controlled Baseline Regimes

Model ReleasesDGX agent

arXiv:2605.23237v1 Announce Type: new Abstract: Stereo image and video generation, stereo geometry estimation, and condition-controlled view synthesis require paired data in which the variables that d

TCAP: Tri-Component Attention Profiling for Unsupervised Backdoor Detection in MLLM Fine-Tuning

ResearchDGX agent

arXiv:2601.21692v2 Announce Type: replace Abstract: Fine-Tuning-as-a-Service (FTaaS) facilitates the customization of Multimodal Large Language Models (MLLMs) but introduces critical backdoor risks vi

The Deterministic Horizon: Impossibility Results as Design Specifications for Trustworthy AI Systems

ApplicationsDGX agent

arXiv:2605.23024v1 Announce Type: new Abstract: Large language models now write software, draft legal documents, and produce clinical notes, yet fundamental limits, from Turing and Arrow to the No Fre

Vector Retrieval with Similarity and Diversity: How Hard Is It?

Model ReleasesDGX agent

arXiv:2407.04573v4 Announce Type: replace-cross Abstract: Dense vector retrieval is an important building block of modern machine learning systems, underlying applications ranging from semantic search

VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction

SafetyDGX agent

arXiv:2602.12579v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a dominant paradigm for enhancing Large Language Models (LLMs) reasoning,

What Training Data Teaches RL Memory Agents: An Empirical Study of Curriculum Effects in Memory-Augmented QA

Model ReleasesDGX agent

arXiv:2605.23067v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a viable recipe for training LLM agents to reason over external memory banks in multi-session dialogue. Exist

When Good Equations Get Bad Scores: Improving Symbolic Regression Through Better Parameter Optimization

Model ReleasesDGX agent

arXiv:2605.23272v1 Announce Type: cross Abstract: Symbolic Regression (SR) plays a central role in scientific knowledge discovery by distilling mathematical equations from observational data. Most exi

Wordle 1,801 4/6 🟨🟨⬛⬛⬛ ⬛🟩⬛⬛⬛ ⬛🟩🟩🟨⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result where the player solved puzzle #1,801 in four attempts using color-coded feedback (yellow for correct letters in wrong positions, green for correct letters in

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention

Model ReleasesDGX agent

arXiv:2502.04230v3 Announce Type: replace-cross Abstract: The rapid proliferation of generative audio synthesis and editing technologies has raised serious concerns about copyright infringement, data

24 May 2026

🇺🇸🇪🇺 A new study from the US-based New England Journal of Medicine found that Americans die earlier across all income levels compared to…

Model ReleasesDGX agent

🇺🇸🇪🇺 A new study from the US-based New England Journal of Medicine found that Americans die earlier across all income levels compared to their European counterparts. What’s especially notable is that

Aleph 2.0 will blow your mind

Model ReleasesDGX agent

Aleph 2.0 will blow your mind Just tested Runway Aleph 2.0 and this blew my mind a bit lol I saw a video like this when Aleph first released and with the new Aleph 2.0 I wanted to create my own versio

An interesting work on Physical AI: PhysX-Omni. First unified sim-ready generation framework for rigid, deformable, and articulated objects,…

Model ReleasesDGX agent

An interesting work on Physical AI: PhysX-Omni. First unified sim-ready generation framework for rigid, deformable, and articulated objects, with a diverse dataset and new benchmark. 🌐 https://physx-o

Built an AI screen memory using llama.cpp + Gemma 4 — remembers everything you do on your computer,search/chat or make agents over it. 100% local

Model ReleasesDGX agent

This project demonstrates a local AI system built with llama.cpp and Gemma 4 that captures and analyzes screen activity to create persistent memory of user computer interactions, enabling search, chat

It makes many online spaces intolerable. If I want to talk to ChatGPT or Claude, I'll just talk to ChatGPT or Claude, I don't need to talk t…

Model ReleasesDGX agent

It makes many online spaces intolerable. If I want to talk to ChatGPT or Claude, I'll just talk to ChatGPT or Claude, I don't need to talk to ChatGPT and Claude pretending to be DoofWarrior123 on X wi

It’s no longer just AI companies & their founders being sued over AI training - individual researchers are now being sued, too. In a new law…

Model ReleasesDGX agent

It’s no longer just AI companies & their founders being sued over AI training - individual researchers are now being sued, too. In a new lawsuit, two authors allege that Guillaume Lample, while an AI

Mad House — Usborne Creepy Computer Games

Model ReleasesDGX agent

Tool: Mad House — Usborne Creepy Computer Games Via Hacker News I learned that UK publisher Usborne published free PDFs of their 1980s Computer Books, some of which I remember working through on my Co

People often ask what my biggest tip is for getting the most out of Claude Code. These days my #1 tip is: use auto mode Auto mode means no m…

Model ReleasesDGX agent

People often ask what my biggest tip is for getting the most out of Claude Code. These days my #1 tip is: use auto mode Auto mode means no more permission prompts. It is the key building block for mul

quick summary of someone's github, cool! here's me

Model ReleasesDGX agent

quick summary of someone's github, cool! here's me I always wanted a GitHub dashboard: See my repos, open Issues/PRs, what version I released last, how many commits since last release. So I built one

Wordle 1,799 4/6 ⬛⬛⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛🟨⬛⬛⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post shows a Wordle game result where the player solved puzzle #1,799 in 4 attempts, with the final answer being a five-letter word where all letters are in the correct positions (indicated by th

Wordle 1,800 4/6 ⬛⬛⬛⬛🟩 ⬛🟩🟨⬛⬛ 🟩🟩🟨⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result where the player solved puzzle #1,800 in 4 attempts, using the color-coded feedback system (gray for wrong letters, yellow for correct letters in wrong positio

23 May 2026

1. Agreed w @scaling01 that Mythos appears to be better GPT 5.5 on many metrics. 2. Mythos is definitely a major wakeup call wrt security, a…

Model ReleasesDGX agent

1. Agreed w @scaling01 that Mythos appears to be better GPT 5.5 on many metrics. 2. Mythos is definitely a major wakeup call wrt security, and will pose problems for real-world systems that aren’t wel

An entropy formula for the Deep Linear Network

Model ReleasesDGX agent

arXiv:2509.09088v3 Announce Type: replace Abstract: We study the Riemannian geometry of the Deep Linear Network (DLN) as a foundation for a thermodynamic description of the learning process. The main

An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning

Model ReleasesDGX agent

arXiv:2605.21968v1 Announce Type: new Abstract: Optimization is essential in deep learning. The foundational method upon which most optimizers are built is momentum-based stochastic gradient descent.

ASSEMBLAGE-DEEPHISTORY: A Cross-Build Binary Dataset with Temporal Coverage

Model ReleasesDGX agent

arXiv:2605.21615v1 Announce Type: cross Abstract: Existing binary corpora typically capture only one or two axes of binary variation: they either provide cross-compiler builds without a temporal axis,

b9296

Local AiDGX agent

b9296 is a build release of llama.cpp , the C/C++ implementation for efficient large language model inference. As an intermediate build in the llama.cpp release cycle, it likely includes recent bug fi

Chebyshev Policies and the Mountain Car Problem: Reinforcement Learning for Low-Dimensional Control Tasks

Model ReleasesDGX agent

arXiv:2605.22305v1 Announce Type: new Abstract: We analytically solve the Mountain Car problem, a canonical benchmark in RL, and derive an optimal control solution, closing a gap after 36 years. This

co-sign. a very handy mental framework for what kinds of learning transformers do well today, and why it runs into limitations. when @ankit2…

ToolsDGX agent

co-sign. a very handy mental framework for what kinds of learning transformers do well today, and why it runs into limitations. when @ankit2119 and i wrote about the need for adversarial world models

Cross-domain benchmarks reveal when coordinated AI agents improve scientific inference from partial evidence

Model ReleasesDGX agent

arXiv:2605.22300v1 Announce Type: cross Abstract: Scientific evidence often spans instruments, databases, and disciplines, so no single source records the full phenomenon. This makes it difficult to d

Cross-Species RSA Reveals Conserved Early Visual Alignment but Divergent Higher-Area Rankings Across Human fMRI and Macaque Electrophysiology

SafetyDGX agent

arXiv:2605.22401v1 Announce Type: new Abstract: Does the relationship between learning rules and brain alignment generalize across species? We extend our prior finding that untrained CNNs match backpr

Do Not Trust The Auctioneer: Learning to Bid in Feedback-Manipulated Auctions

Model ReleasesDGX agent

arXiv:2605.22438v1 Announce Type: cross Abstract: Shilling is the use of artificial bids to make competition appear stronger and push prices upward. We study repeated first-price auctions in which shi

Double descent for least-squares interpolation on contaminated data: A simulation study

ResearchDGX agent

arXiv:2605.21494v1 Announce Type: new Abstract: Overparametrized models can exhibit an excellent generalization performance, although they should be prone to overfitting according to classical statist

Dropout Universality: Scaling Laws and Optimal Scheduling at the Edge-of-Chaos

Model ReleasesDGX agent

arXiv:2605.21648v1 Announce Type: new Abstract: We develop a mean-field theory of dropout as a perturbation of critical signal propagation at the edge of chaos. Dropout shifts the perfect-alignment fi

Dynamic Mixture of Latent Memories for Self-Evolving Agents

ResearchDGX agent

arXiv:2605.21951v1 Announce Type: new Abstract: Achieving self-evolution in intelligent agents requires the continual accumulation of new knowledge across changing task sequences without forgetting pr

EmoTrack: Robust Depression Tracking from Counseling Transcripts across Session Regimes

Model ReleasesDGX agent

arXiv:2605.22286v1 Announce Type: new Abstract: Text-based counseling is an important interface for AI mental-health support, where transcripts may be used to monitor depression severity and flag sess

Evaluation of Pipelines for Data Integration into Knowledge Graphs

Model ReleasesDGX agent

arXiv:2605.22304v1 Announce Type: cross Abstract: Integrating new data into knowledge graphs (KG) typically involves different tasks that are executed within workflows or pipelines There are many poss

@GaryMarcus Can't agree more. OpenAI seems struggling to build any successful product after ChatGPT. Codex is promising but it's still early…

Model ReleasesDGX agent

@GaryMarcus Can't agree more. OpenAI seems struggling to build any successful product after ChatGPT. Codex is promising but it's still early, and it's following Claude Code. It's questionable whether

GPT-5.5 Pro is a very solid fact checker. I can throw entire chapters at it and it will hunt down every key reference accurately. The only r…

Model ReleasesDGX agent

GPT-5.5 Pro is a very solid fact checker. I can throw entire chapters at it and it will hunt down every key reference accurately. The only real annoyance is that it loves nuance, so returns a lot of “

How Anthropic's ongoing discussions with the Vatican about ethics and AI led to Christopher Olah being invited to Pope Leo's unveiling of an encyclical on AI (Jack Jenkins/RNS)

Model ReleasesDGX agent

Jack Jenkins / RNS: How Anthropic's ongoing discussions with the Vatican about ethics and AI led to Christopher Olah being invited to Pope Leo's unveiling of an encyclical on AI — (RNS) — Pope Leo XIV

Hugging Face just released @LeRobotHF a humanoid robot you can build for roughly $2,500! A full open stack with hardware, simulation, traini…

Model ReleasesDGX agent

Hugging Face just released @LeRobotHF a humanoid robot you can build for roughly $2,500! A full open stack with hardware, simulation, training environments, runtime tools, datasets and robot learning

I called both missteps and that’s part of why OpenAI and their shills hate me. 🤷‍♂️

Model ReleasesDGX agent

I called both missteps and that’s part of why OpenAI and their shills hate me. 🤷‍♂️ @GaryMarcus It all went wrong with GPT-5 and Sora 2 IMO. Until then OpenAI could do no wrong in the eyes of most peo

← Previous
1…668669670671672…1042
Next →