AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlog
86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
Safety

ReasonLight: A Multimodal Foundation Model-Enhanced Reinforcement Learning Framework for Zero-Shot Traffic Signal Control

DGX agent

arXiv:2605.29425v1 Announce Type: new Abstract: Reinforcement learning (RL) has shown promise in traffic signal control (TSC). However, its reliance on predefined states limits responsiveness to obser

safetyarxiv-cs-ai
29 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

SAHG: Sector-Anisotropic Hyperbolic Graph Model for Social Bot Detection

DGX agent

arXiv:2605.30166v1 Announce Type: cross Abstract: LLM-driven social bots can generate fluent, human-like text, reducing the discriminative advantage of content-based detection alone. However, coordina

safetyarxiv-cs-lg
29 May 2026
Safety

SigmaMedStat: Temporal Signal Modeling for ICU False Alarm Reduction

DGX agent

arXiv:2605.29236v1 Announce Type: new Abstract: Alarm fatigue in intensive care units (ICUs) is a well documented patient safety crisis. Clinical monitors generate 350 or more alarms per patient per d

safetyarxiv-cs-lg
29 May 2026
Model Releases

Small Agent Group is the Future of Digital Health

DGX agent

arXiv:2602.08013v2 Announce Type: replace Abstract: The rapid adoption of large language models (LLMs) in digital health has been driven by a 'scaling-first' philosophy, i.e., the assumption that clin

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

The Trust Paradox: How CS Researchers Engage LLM Leaderboards

DGX agent

arXiv:2605.28966v1 Announce Type: new Abstract: Large language model (LLM) leaderboards rank AI models using standardized benchmarks and have become highly visible across computer science, despite kno

model-releasesarxiv-cs-cl
29 May 2026
Research

Towards Continuous-time Causal Foundation Models

DGX agent

arXiv:2605.28880v1 Announce Type: new Abstract: Extending discrete-time causal Prior-data Fitted Networks for time series to continuous time invites writing the mechanism as a stochastic differential

researcharxiv-cs-lg
29 May 2026
Model Releases

AdaDPO: Self-Adaptive Direct Preference Optimization with Balanced Gradient Updates

DGX agent

arXiv:2605.28440v1 Announce Type: new Abstract: DPO has become a widely adopted alternative to RLHF for aligning LLMs with human preferences, eliminating the need for a separate reward model or RL loo

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

AI in SRE: Where and how Google is deploying agentic AI to improve operations

DGX agent

Since its inception over 20 years ago, Google has used Site Reliability Engineering (SRE) to keep services like Search, Gmail, Maps, YouTube and Google Cloud reliable and highly available, adhering to

model-releasesgoogle-cloud-ai
28 May 2026
Research

Applications of temporal graph learning for predicting the dynamics of biological systems

DGX agent

arXiv:2605.28659v1 Announce Type: new Abstract: Biological foundation models have shown strong performance in single-cell representation learning by applying transformer architectures directly to gene

researcharxiv-cs-lg
28 May 2026
Model Releases

Claude Opus 4.8: 'a modest but tangible improvement'

DGX agent

Anthropic shipped Claude Opus 4.8 today. My favourite thing about it is this note in the release announcement: Users will find Opus 4.8 to be a modest but tangible improvement on its predecessor. Ther

model-releasessimon-willison
28 May 2026
Research

Debate Helps Weak Judges Reward Stronger Models

DGX agent

arXiv:2605.27483v1 Announce Type: cross Abstract: Despite theoretical promise, debate as a scalable oversight protocol has produced mixed empirical results: gains in some settings, and null effects in

researcharxiv-cs-ai
28 May 2026
Safety

DebFilter: Eradicating Biases Stashed in Value

DGX agent

arXiv:2605.28167v1 Announce Type: new Abstract: Text-to-image diffusion models, which are theoretically equivalent to score-based generative models, generate images through a multi-step denoising proc

safetyarxiv-cs-cv
28 May 2026
Model Releases

FPMoE: A Sparse Mixture-of-Experts Approach to Functional Code Generation

DGX agent

arXiv:2605.27849v1 Announce Type: cross Abstract: Despite rapid progress in LLM-based code generation, existing models are predominantly trained on imperative languages, leaving functional programming

model-releasesarxiv-cs-ai
28 May 2026
Agents

Heterogeneous Multi-Agent Modeling for Measurement and Network Analysis of the Data Service Market

DGX agent

arXiv:2605.27433v1 Announce Type: cross Abstract: With the increasing complexity of collaboration among various social entities and user demands, the factors affecting the stable development of the da

agentsarxiv-cs-ai
28 May 2026
Research

Hybrid Neural World Models

DGX agent

arXiv:2605.28317v1 Announce Type: cross Abstract: Neural surrogates promise large speedups over classical solvers for physical dynamics but fail silently at sharp dynamical events such as shocks, fron

researcharxiv-cs-ai
28 May 2026
Model Releases

KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs

DGX agent

arXiv:2605.27984v1 Announce Type: cross Abstract: Speech language models (SpeechLMs) have achieved substantial progress by extending large language models (LLMs) to the speech modality. However, Speec

model-releasesarxiv-cs-ai
28 May 2026
Tutorials

Learning the Error Patterns of Language Models

DGX agent

arXiv:2605.28328v1 Announce Type: cross Abstract: When generating outputs for domains with specific validity constraints (e.g., a program should compile), LLMs often fail in a small number of focused

tutorialsarxiv-cs-ai
28 May 2026
Safety

Mathematical Modelling of Ethical AI Use in Higher Education: A Coordination Game Framework for Future-Facing Learning

DGX agent

arXiv:2605.27400v1 Announce Type: cross Abstract: The rapid uptake of generative artificial intelligence (AI) in higher education is reshaping assessment practices and intensifying concerns around aca

safetyarxiv-cs-ai
28 May 2026
Applications

Measuring Massive Multitask Chinese Understanding

DGX agent

arXiv:2304.12986v3 Announce Type: replace-cross Abstract: The development of large-scale Chinese language models is flourishing, yet there is a lack of corresponding capability assessments. Therefore,

applicationsarxiv-cs-ai
28 May 2026
Model Releases

MTAVG-Bench 2.0: Diagnosing Failure Modes of Cinematic Expressiveness in Multi-Talker Audio-Video Generation

DGX agent

arXiv:2605.28035v1 Announce Type: new Abstract: In recent years, Multi-Talker Audio-Video Generation (MTAVG) models have shown promising performance on fundamental metrics such as lip-sync and audio-v

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Multi-Adapter Representation Interventions via Energy Calibration

DGX agent

arXiv:2605.28722v1 Announce Type: new Abstract: Representation intervention has emerged as a promising paradigm for aligning large language models toward desired behaviors without modifying model weig

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

MUSE: Benchmarking Manufacturable, Functional, and Assemblable Text-to-CAD Generation

DGX agent

arXiv:2605.28579v1 Announce Type: new Abstract: Large language models (LLMs) have recently advanced text-driven 3D generation, yet Text-to-CAD remains far from supporting industrial product design. Ex

model-releasesarxiv-cs-ai
28 May 2026
Applications

SMILE-Next: Teaching Large Language Models to Detect, Classify, and Reason about Laughter

DGX agent

arXiv:2605.28084v1 Announce Type: cross Abstract: Laughter is a complex social signal that conveys communicative intent beyond amusement. While prior work has focused on isolated laughter analysis tas

applicationsarxiv-cs-ai
28 May 2026
Model Releases

The Alignment Floor: When Persona Customization Is Safe

DGX agent

arXiv:2605.27382v1 Announce Type: cross Abstract: A key promise of pluralistic AI is behavioral adaptation: persona prompts like 'be creative' or 'be thorough' let systems respect diverse user values

model-releasesarxiv-cs-ai
28 May 2026
Safety

TRACES: Proactive Safety Auditing for Multi-Turn LLM Agents via Trajectory-State Modeling

DGX agent

arXiv:2605.27690v1 Announce Type: new Abstract: LLM agents increasingly operate through multi-turn tool use and environment interaction, where safety risks often emerge from intermediate steps long be

safetyarxiv-cs-cl
28 May 2026
Model Releases

ADRD-Bench: A Preliminary LLM Benchmark for Alzheimer's Disease and Related Dementias

DGX agent

arXiv:2602.11460v2 Announce Type: replace Abstract: Large language models (LLMs) have shown great potential for healthcare applications. However, existing evaluation benchmarks provide minimal coverag

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Benchmark Leakage Trap: Can We Trust LLM-based Recommendation?

DGX agent

arXiv:2602.13626v3 Announce Type: replace Abstract: The expanding integration of Large Language Models (LLMs) into recommender systems poses critical challenges to evaluation reliability. This paper i

model-releasesarxiv-cs-lg
27 May 2026
Research

Boosting Knowledge Graph Foundation Models via Enhanced Negative Sampling

DGX agent

arXiv:2605.27023v1 Announce Type: new Abstract: Knowledge graphs (KGs) have become the core backbone of numerous downstream tasks such as question answering and recommender systems. However, despite a

researcharxiv-cs-ai
27 May 2026
Applications

Conceptual Schema Inference for Tabular Datasets using Large Language Models

DGX agent

arXiv:2509.04632v2 Announce Type: replace-cross Abstract: Large collections of tabular data from data lakes, web tables and open data portals often originate from heterogeneous sources, leading to rep

applicationsarxiv-cs-ai
27 May 2026
Model Releases

Constraint acquisition needs better benchmarks

DGX agent

arXiv:2605.26279v1 Announce Type: new Abstract: Constraint Acquisition (CA) and related research on the validation and enhancement of Mathematical Programming (MP) models from domain knowledge artifac

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

DunbaaBERT: From Sacrifice to Semantics

DGX agent

arXiv:2605.26935v1 Announce Type: new Abstract: Large language models have achieved strong performance across many NLP tasks, yet Urdu remains comparatively underexplored due to limited resources and

model-releasesarxiv-cs-cl
27 May 2026
Safety

EmoDistill: Offline Emotion Skill Distillation for Language Model Agents in Adversarial Negotiation

DGX agent

arXiv:2605.26785v1 Announce Type: cross Abstract: Post-trained LLMs are often optimized to align responses with human preferences, making them safe, polite, and conversationally appropriate. In advers

safetyarxiv-cs-ai
27 May 2026
Agents

FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentation

DGX agent

arXiv:2605.27178v1 Announce Type: cross Abstract: We address the challenging task of 3D object segmentation in complex scene point clouds without relying on any scene-level human annotations during tr

agentsarxiv-cs-ai
27 May 2026
Model Releases

Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction

DGX agent

arXiv:2605.26230v1 Announce Type: new Abstract: Multi-view 3D reconstruction has achieved remarkable progress with the advent of feed-forward 3D reconstruction models. However, these models are typica

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

I think Anthropic and OpenAI have found product-market fit

DGX agent

Anthropic are strongly rumored to be about to have their first profitable quarter. Stories are circulating of companies surprised at how expensive their LLM bills are becoming from usage by their staf

model-releasessimon-willison
27 May 2026
Research

MedGuideX: Internalizing Decision Logic from Executable Guidelines into Large Language Models for Clinical Reasoning

DGX agent

arXiv:2605.26567v1 Announce Type: new Abstract: Clinical practice guidelines (CPGs) encode evidence-based decision logic that clinicians apply by evaluating patient variables, conditional criteria, an

researcharxiv-cs-ai
27 May 2026
Model Releases

MTL-FNO: A Lightweight Multi-Task Fourier Neural Operator for Sparse Field Reconstruction

DGX agent

arXiv:2605.26718v1 Announce Type: new Abstract: Efficient onboard multi-field sparse reconstruction is essential for the autonomous operation of aerospace vehicles. While existing deep learning models

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

Prompt Injection Detection is Regime-Dependent: A Deployment-Aware Evaluation with Interpretable Structural Signals

DGX agent

arXiv:2605.26999v1 Announce Type: new Abstract: Prompt injection poses a critical threat to the safe deployment of large language models, yet existing detection approaches are typically evaluated unde

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Stability Implies Redundancy: Delta Attention Selective Halting for Efficient Long-Context Prefilling

DGX agent

arXiv:2604.18103v2 Announce Type: replace Abstract: Prefilling computational costs pose a significant bottleneck for Large Language Models (LLMs) and Large Multimodal Models (LMMs) in long-context set

model-releasesarxiv-cs-ai
27 May 2026
Hardware

Tensormesh taps Nvidia, AMD and CoreWeave for funding to fix AI model memory problems

DGX agent

Tensormesh Inc. has hit upon a way to make artificial intelligence inference more efficient by eliminating the need for redundant computations, and its technology is so convincing that several of AI i

hardwaresiliconangle
27 May 2026
Model Releases

The Bridge-Garden Dilemma in LLM Distillation: Why Mixing Hard and Soft Labels Works

DGX agent

arXiv:2605.26246v1 Announce Type: new Abstract: Knowledge distillation (KD) transfers knowledge from a large teacher model to a smaller student. In language modeling, the student is trained either on

model-releasesarxiv-cs-lg
27 May 2026
Safety

When Does LeJEPA Learn a World Model?

DGX agent

arXiv:2605.26379v1 Announce Type: cross Abstract: A representation that scrambles the true degrees of freedom of the world cannot support reliable planning or compositional generalization. We prove th

safetyarxiv-cs-lg
27 May 2026
Model Releases

When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation

DGX agent

arXiv:2509.26600v2 Announce Type: replace-cross Abstract: As LLMs rapidly saturate existing benchmarks, automated benchmark creation using LLMs (LLM-as-a-benchmark) -- where a model generates test inp

model-releasesarxiv-cs-ai
27 May 2026
Agents

Act or Clarify? Modeling Sensitivity to Uncertainty and Cost in Communication

DGX agent

arXiv:2602.02843v3 Announce Type: replace Abstract: When deciding how to act under uncertainty, agents may choose to act to reduce uncertainty or they may act despite that uncertainty. In communicativ

agentsarxiv-cs-cl
26 May 2026
Model Releases

An Efficient Learning Method to Connect Observables

DGX agent

arXiv:2503.01684v3 Announce Type: replace-cross Abstract: Constructing fast and accurate surrogate models is a key ingredient for making robust predictions in many topics. We introduce a new model, th

model-releasesarxiv-cs-lg
26 May 2026
Safety

Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation

DGX agent

arXiv:2605.25402v1 Announce Type: cross Abstract: Self-supervised pre-training paradigm has gained increasing prominence for learning transferable representations in medical imaging, yet existing meth

safetyarxiv-cs-ai
26 May 2026
Agents

APT-Agent: Automated Penetration Testing using Large Language Models

DGX agent

arXiv:2605.24949v1 Announce Type: cross Abstract: Penetration testing is essential to securing modern web infrastructures, yet traditional manual methods struggle to keep pace with their scale and com

agentsarxiv-cs-ai
26 May 2026
Model Releases

Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction

DGX agent

arXiv:2605.24657v1 Announce Type: new Abstract: Major LLM platforms deploy models in an inference-only configuration: the model serves requests but never updates per-user weights. Users must repeatedl

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…326327328329330…1302
Next →