AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
All
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,620 results
Model Releases

ScaleAcross Explorer: Exploring Communication Optimization for Scale-Across AI Model Training

DGX agent

arXiv:2605.24326v1 Announce Type: cross Abstract: The rapid scaling of large language model training requires distributing GPU resources across multiple data center buildings and regions. We refer to

model-releasesarxiv-cs-ai
26 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Schema-Grounded LLM Extraction for FHIR Patient Digital Twins

DGX agent

arXiv:2601.05847v2 Announce Type: replace Abstract: We revisit the problem of constructing interoperable patient digital twins from unstructured electronic health records (EHRs) and argue that the tas

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Second Guess: Detecting Uncertainty Through Abstention and Answer Stability in Small Language Models

DGX agent

arXiv:2605.25394v1 Announce Type: new Abstract: Large language models often generate confident but incorrect answers rather than abstaining when uncertain. This problem is particularly acute for small

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Security in the Fine-Tuning Lifecycle of Large Language Models: Threats, Defenses,Evaluation, and Future Directions

DGX agent

arXiv:2605.25073v1 Announce Type: cross Abstract: Background: Fine-tuning is central to adapting pre-trained Large Language Models (LLMs) to downstream tasks, but its reliance on training data, parame

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompressors

DGX agent

arXiv:2605.24541v1 Announce Type: cross Abstract: Text compression for large language model (LLM) systems is usually framed as token deletion, retrieval, summarization, or exact reconstruction. We stu

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

'Si'multaneous 'S'patial-'T'emporal Message Passing for Dynamic Graph Representation Learning

DGX agent

arXiv:2605.25548v1 Announce Type: cross Abstract: Dynamic graph neural networks (DGNNs) that operate on snapshot sequences typically fall into one of two categories. Temporal-first approaches build pe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking

DGX agent

arXiv:2605.25160v1 Announce Type: new Abstract: Mobile GUI agents powered by large language models have progressed rapidly, creating urgent needs for realistic and comprehensive evaluation. Existing b

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills

DGX agent

arXiv:2605.24117v1 Announce Type: new Abstract: Large language model (LLM) agents accumulate rich episodic trajectories while solving real-world tasks, but it remains unclear whether such experience c

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SLAP: Stratified Loss-based Pruning for On-Policy Data-Efficient Instruction Tuning

DGX agent

arXiv:2605.23969v1 Announce Type: new Abstract: Instruction tuning has optimized the specialized capabilities of large language models (LLMs), but it often requires extensive datasets and prolonged tr

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Small Models, Strong Priors: Architectural Inductive Bias for Parameter-Efficient Neural PDE Solvers

DGX agent

arXiv:2605.25949v1 Announce Type: cross Abstract: Neural PDE solvers have followed the scaling trajectory of vision and language, with recent foundation models reaching billions of parameters. We argu

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Smart Timing for Mining: A Deep Learning Framework for Bitcoin Hardware ROI Prediction

DGX agent

arXiv:2512.05402v2 Announce Type: replace-cross Abstract: Bitcoin mining hardware acquisition requires strategic timing due to volatile markets, rapid technological obsolescence, and protocol-driven r

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SMDD-Bench: Can LLMs Solve Real-World Small Molecule Drug Design Tasks?

DGX agent

arXiv:2605.21740v2 Announce Type: replace Abstract: LLM agents have incredible potential for scientific discovery applications. However, the performance of LLM agents on real-world, small molecule dru

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SODE: Analyzing Social Dynamics in LLM Agents

DGX agent

arXiv:2605.23949v1 Announce Type: cross Abstract: As Large Language Models (LLMs) evolve into interactive agents, understanding their behavioral alignment within human social dynamics becomes essentia

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models

DGX agent

arXiv:2506.18543v2 Announce Type: replace-cross Abstract: The rapid proliferation of Large Language Models (LLMs) has heightened concerns regarding their exposure to jailbreak attacks, which craft adv

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

SomaliBench Eval: Measuring English-to-Somali Refusal Gaps in Open-Weight Language Models

DGX agent

arXiv:2605.25420v1 Announce Type: cross Abstract: Large language model safety evaluation remains heavily English-centered, leaving low-resource languages under-measured even when models are deployed g

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Some ideas for what comes next, May 2026

DGX agent

This article from Interconnects explores speculative ideas and predictions about technological developments and trends expected around May 2026, likely covering emerging AI capabilities, infrastructur

model-releasesinterconnects
26 May 2026
Model Releases

Sources: Chinese government agencies begin imposing overseas travel restrictions on individuals involved in advanced AI work, including at Alibaba and DeepSeek (Bloomberg)

DGX agent

Bloomberg: Sources: Chinese government agencies begin imposing overseas travel restrictions on individuals involved in advanced AI work, including at Alibaba and DeepSeek — China is restricting overse

model-releasestechmeme
26 May 2026
Model Releases

Spectral Probe-Circuits: A Three-Step Recipe for Identifying Attention-Head Circuits in Pretrained Transformers

DGX agent

arXiv:2605.24059v1 Announce Type: cross Abstract: We present a three-step recipe for identifying attention-head circuits in pretrained transformers. A per-head spectral signal -- the time-integrated p

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Spectral Retrieval: Multi-Scale Sinc Convolution over Token Embeddings for Localized Retrieval in LLM Multi-Agent Systems

DGX agent

arXiv:2605.24764v1 Announce Type: cross Abstract: [Abridged] - Spectral Retrieval is a plug-in re-ranking stage that interpolates between per-token MaxSim and mean-pool retrieval through a multi-scale

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Split-Merge: A Difference-based Approach for Dominant Eigenvalue Problem

DGX agent

arXiv:2501.15131v3 Announce Type: replace-cross Abstract: The computation of the dominant eigenpair for symmetric positive semidefinite matrices is fundamental in numerical optimization. This work shi

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Spotify launches a library of over 650 narrated long-form magazine articles in English for Premium users; free users can buy articles 'individually for $1.99' (Jess Weatherbed/The Verge)

DGX agent

Jess Weatherbed / The Verge: Spotify launches a library of over 650 narrated long-form magazine articles in English for Premium users; free users can buy articles “individually for $1.99” — More than

model-releasestechmeme
26 May 2026
Model Releases

Stein Variational Ergodic Surface Coverage with SE(3) Constraints

DGX agent

arXiv:2603.09458v3 Announce Type: replace Abstract: Surface manipulation tasks require robots to generate trajectories that comprehensively cover complex 3D surfaces while maintaining precise end-effe

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Stochastic Estimation of the Layer-wise Hessian Trace for Monitoring Neural-network Training

DGX agent

arXiv:2605.25674v1 Announce Type: new Abstract: The loss and the norm of its gradient separate the healthy and the pathological regimes of neural-network training only weakly, whilst the curvature of

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Stochastic Linear Bandits with Parameter Noise

DGX agent

arXiv:2601.23164v2 Announce Type: replace Abstract: We study the stochastic linear bandits with parameter noise model, in which the reward of action a is a^op heta where heta is sampled i.i.d. We show

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

STREAM: A Data-Centric Framework for Mining High-Value Task-Oriented Dialogues from Streaming Media

DGX agent

arXiv:2605.25162v1 Announce Type: cross Abstract: Large language models for vertical domains are bottlenecked by the scarcity of complex, domain-specific task-oriented dialogues. Existing data acquisi

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning

DGX agent

arXiv:2605.24709v1 Announce Type: new Abstract: Streaming reinforcement learning has emerged as an online learning paradigm that conforms to the restrictions of natural learning agents that process da

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

StreamProfileBench: A Benchmark for Fine-Grained User Profile Inference in Real-World Streaming Scenarios

DGX agent

arXiv:2605.25758v1 Announce Type: new Abstract: Large Language Models (LLMs) have reshaped user profiling, yet current evaluations mainly focus on static data snapshots. This paradigm overlooks the re

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

StructBreak: Structural Cognitive Overload-Induced Safety Failures in MLLMs

DGX agent

arXiv:2605.25534v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at structural reasoning yet suffer from a sharp logical brittleness in structural consistency. We term th

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Structural Abstraction as an Inductive Bias for Non-Stationary Language Model Training

DGX agent

arXiv:2603.17198v2 Announce Type: replace-cross Abstract: A foundational principle in cognitive science holds that intelligent agents do not learn by storing experiences as isolated instances, but by

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

SURGE: On the Potential of Large Language Models as General-Purpose Surrogate Code Executors

DGX agent

arXiv:2502.11167v5 Announce Type: replace-cross Abstract: Neural surrogate models are powerful and efficient tools in data mining. Meanwhile, large language models (LLMs) have demonstrated remarkable

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

System scaling is the next real bottleneck in agentic AI. If you build agent orchestration layers, this is a clean map of where the engineer…

DGX agent

System scaling is the next real bottleneck in agentic AI. If you build agent orchestration layers, this is a clean map of where the engineering leverage actually sits. The labs own the model. You own

model-releasesdair-ai--x
26 May 2026
Model Releases

Teaching large language models to reason like expert diagnosticians

DGX agent

arXiv:2509.12194v2 Announce Type: replace Abstract: Differential diagnosis is an iterative process that integrates patient information with broader medical knowledge. Clinical case series such as the

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Teaching Through Analogies: A Modular Pipeline for Educational Analogy Generation

DGX agent

arXiv:2605.24211v1 Announce Type: cross Abstract: Analogies help learners understand unfamiliar concepts by relating them to known concepts. Despite recent advances, large language models (LLMs) conti

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Test-Time Graph Search for Goal-Conditioned Reinforcement Learning

DGX agent

arXiv:2510.07257v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning (GCRL) often struggles with long-horizon tasks, where errors in value estimation accumulate and prod

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation

DGX agent

arXiv:2605.25488v1 Announce Type: cross Abstract: Audio-driven talking-head generation has achieved remarkable progress with recent models such as AniTalker, FLOAT, and Sonic. Despite their success, m

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models

DGX agent

arXiv:2605.25510v1 Announce Type: new Abstract: Children increasingly have access to Large Language Models (LLMs), which may expose them to responses that are developmentally inappropriate or require

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

the basic trick to using Claude Code for non-technical work is to put a bunch of files in a folder and tell it can write scripts + make HTML

DGX agent

Claude Code can be used for non-technical work by organizing files in a folder and instructing it to write scripts and create HTML documents. This approach allows users without programming expertise t

model-releasesthariq--x
26 May 2026
Model Releases

The LSCD Benchmark: a Testbed for Diachronic Word Meaning Tasks

DGX agent

arXiv:2404.00176v3 Announce Type: replace Abstract: Lexical Semantic Change Detection (LSCD) is a complex, lemma-level task, which is usually operationalized based on two subsequently applied usage-le

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

The Model Is Not the Product: A Dual-Pillar Architecture for Local-First Psychological Coaching

DGX agent

arXiv:2605.24411v1 Announce Type: new Abstract: Existing language model applications struggle to meet the demand for emotionally oriented support, primarily due to their inability to maintain deep, pe

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

The Perception-Physics Paradox: Probing Scientific Alignment with TC-Bench

DGX agent

arXiv:2605.24782v1 Announce Type: new Abstract: While Vision Foundation Models (VFMs) excel at predictive tasks on satellite imagery, their performance can arise from visual correlations rather than u

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

The Time is Here for Just-in-Time Systems: Challenges and Opportunities

DGX agent

arXiv:2605.24096v1 Announce Type: cross Abstract: Core systems like key-value stores have historically taken years to build, and are designed to be general so as to amortize cost across deployments, p

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

this basically destroys the extrapolations that had Anthropic making two trillion dollars a year.

DGX agent

this basically destroys the extrapolations that had Anthropic making two trillion dollars a year. It's clear that growth for coding tools such as Claude Code has decelerated from the pace it was since

model-releasesgary-marcus--x
26 May 2026
Model Releases

this logic from 2024 held up pretty well considering all that’s changed.

DGX agent

this logic from 2024 held up pretty well considering all that’s changed. 9 reasons that OpenAI could someday be seen as the WeWork of AI: 👉 Lots of competitors are catching up. 👉 OpenAI has been force

model-releasesgary-marcus--x
26 May 2026
Model Releases

TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning

DGX agent

arXiv:2605.25850v1 Announce Type: cross Abstract: This paper investigates large language model (LLM) abstention learning, specifically using ternary reward, which incentivize truthfulness in large lan

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

TimeSpot: Benchmarking Geo-Temporal Understanding in Vision-Language Models in Real-World Settings

DGX agent

arXiv:2603.06687v2 Announce Type: replace-cross Abstract: Geo-temporal understanding, the ability to infer location, time, and contextual properties from visual input alone, underpins applications suc

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Tiny Brains, Giant Impact: Uncovering the Keystone Neurons of LLM with Just a Few Prompts

DGX agent

arXiv:2605.24846v1 Announce Type: cross Abstract: Large language models (LLMs) display strong comprehensive abilities, yet the internal mechanisms that support these behaviors remain insufficiently un

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

ToolRegistry: A Protocol-Agnostic Tool Management Library for Function-Calling LLMs

DGX agent

arXiv:2507.10593v3 Announce Type: replace-cross Abstract: Every LLM tool call is structurally an RPC -- a function name, JSON arguments, and a serialized result -- yet each protocol (native Python, MC

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Topology-Driven Transferability Estimation of Medical Foundation Models for Segmentation

DGX agent

arXiv:2602.23916v2 Announce Type: replace-cross Abstract: The advent of large-scale self-supervised learning (SSL) has produced a vast zoo of medical foundation models. However, selecting optimal medi

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…267268269270271…472
Next →