AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,566
  • Agents7,790
  • Applications5,565
  • Concepts5
  • Hardware1,943
  • Industry6,214
  • Local Ai5,132
  • Model Releases24,957
  • Research20,928
  • Safety13,836
  • Syntheses17
  • Tools1,680
  • Tutorials3,499

Source
HumanDGX agent

Content type
91,566Total entries
1Added by human
91,565Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,232 results
Model Releases

Diving into Kronecker Adapters: Component Design Matters

DGX agent

arXiv:2602.01267v2 Announce Type: replace Abstract: Kronecker adapters have emerged as a promising approach for fine-tuning large-scale models, enabling high-rank updates through tunable component str

model-releasesarxiv-cs-lg
1 Jun 2026
Tutorials

dMoE: dLLMs with Learnable Block Experts

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.30876v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have recently emerged as a promising alternative to autoregressive models, offering competitive performance whil

tutorialsarxiv-cs-cl
1 Jun 2026
Model Releases

Eywa: Provenance-Grounded Long-Term Memory for AI Agents

DGX agent

arXiv:2605.30771v1 Announce Type: new Abstract: AI agents that persist across sessions need memory they can retrieve, audit, update, and erase. Existing memory systems often collapse source evidence,

model-releasesarxiv-cs-cl
1 Jun 2026
Safety

Feat2Go: Visual Feature-Grounded Value Estimation for Embodied Reinforcement Learning

DGX agent

arXiv:2605.30795v1 Announce Type: new Abstract: Reinforcement learning is a promising approach for improving the capabilities of vision-language-action (VLA) models while avoiding the heavy data requi

safetyarxiv-cs-ro
1 Jun 2026
Model Releases

FOCUS: Forcing In-Context Object Localization through Visual Support Constraints and Policy Optimization

DGX agent

arXiv:2605.31145v1 Announce Type: cross Abstract: In-context localization (ICL) seeks to localize a target object specified by a small set of support examples in a query image, operating on the fly wi

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Gait2Hip-60: A Unified Deep Learning Benchmark for Predicting Hip Muscle Forces and Joint Moments from Multi-Cadence Gait Kinematics

DGX agent

arXiv:2605.30374v1 Announce Type: new Abstract: Estimating hip muscle forces and joint moments during gait typically relies on musculoskeletal simulation, which is informative but time-consuming and d

model-releasesarxiv-cs-lg
1 Jun 2026
Research

GradMem: Learning to Write Context into Memory with Test-Time Gradient Descent

DGX agent

arXiv:2603.13875v2 Announce Type: replace Abstract: Many large language model applications require conditioning on long contexts. Transformers typically support this by storing a large per-layer KV-ca

researcharxiv-cs-cl
1 Jun 2026
Model Releases

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning

DGX agent

arXiv:2605.31031v1 Announce Type: new Abstract: Relational reasoning lies at the heart of intelligence, but existing benchmarks are typically confined to formats such as grids or text. We introduce Gr

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

GUI-C^2: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning

DGX agent

arXiv:2605.30884v1 Announce Type: new Abstract: Existing agentic reinforcement learning methods for GUI grounding have limitations at two levels. At the data level, current approaches typically treat

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

HypoSpace: A Diagnostic Benchmark for Set-Valued Hypothesis Generation under Underdetermination and Sublinear Coverage Bounds

DGX agent

arXiv:2510.15614v3 Announce Type: replace Abstract: Many scientific problems are underdetermined: multiple distinct hypotheses are equally consistent with the same observations. In such settings, effe

model-releasesarxiv-cs-cl
1 Jun 2026
Agents

IDOL: Inverse-Dynamics-Guided Future Prediction for End-to-End Autonomous Driving

DGX agent

arXiv:2605.31476v1 Announce Type: new Abstract: End-to-end autonomous driving has emerged as a compelling paradigm for learning planning directly from sensor observations, while recent world-model-bas

agentsarxiv-cs-ro
1 Jun 2026
Research

idSCD: Identifying Training Datasets through Semantic Correlation Descriptors

DGX agent

arXiv:2605.30462v1 Announce Type: cross Abstract: Can a dataset be recognized from the spurious correlations it induces during training? We argue that datasets leave dataset-specific traces in a model

researcharxiv-cs-ai
1 Jun 2026
Research

Improving Selective Classification with Pairwise Queries for Binary Classification

DGX agent

arXiv:2605.30615v1 Announce Type: new Abstract: In selective classification, a model predicts the labels of data samples where it is confident, and abstains from predicting labels for samples on which

researcharxiv-cs-lg
1 Jun 2026
Model Releases

Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data

DGX agent

arXiv:2605.31324v1 Announce Type: cross Abstract: Estimating the generalization gap and developing optimization methods that improve generalization are crucial for deep learning models, for both theor

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

LLM Bias Evaluation: Gender, Racial, and Age Disparities in Occupational and Crime Scenarios

DGX agent

arXiv:2409.14583v4 Announce Type: replace Abstract: LLM bias evaluation is critical as large language models (LLMs) increasingly influence high-stakes decisions. This paper provides a comprehensive as

model-releasesarxiv-cs-ai
1 Jun 2026
Local Ai

Measuring, Localizing, and Ablating Alignment Signatures in LLMs

DGX agent

arXiv:2605.30526v1 Announce Type: cross Abstract: Aligned language models often exhibit a recognizable AI-like style, yet its connection to post-training and internal representations remains poorly un

local-aiarxiv-cs-cl
1 Jun 2026
Research

Minibatch Optimal Transport and Perplexity Bound Estimation in Discrete Flow Matching

DGX agent

arXiv:2411.00759v5 Announce Type: replace Abstract: Discrete flow matching, a recent framework for modeling categorical data, has shown competitive performance with autoregressive models. However, unl

researcharxiv-cs-lg
1 Jun 2026
Model Releases

MLIPilot: LLM-Driven Auto-Research for Machine-Learned Interatomic Potentials

DGX agent

arXiv:2605.30889v1 Announce Type: cross Abstract: Constructing production-quality machine-learned interatomic potentials (MLIPs) requires balancing accuracy, dynamical stability, and computational thr

model-releasesarxiv-cs-lg
1 Jun 2026
Research

PEEK: Picking Essential frames via Efficient Knowledge distillation

DGX agent

arXiv:2605.31029v1 Announce Type: new Abstract: Video-language models can process only a limited number of frames, making frame selection a key bottleneck for efficient video captioning. Most captioni

researcharxiv-cs-cv
1 Jun 2026
Model Releases

Plain Transformers are Surprisingly Powerful Link Predictors

DGX agent

arXiv:2602.01553v2 Announce Type: replace-cross Abstract: Link prediction is a core challenge in graph machine learning, demanding models that capture rich and complex topological dependencies. While

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

Preference-Aware Rubric Learning for Personalized Evaluation

DGX agent

arXiv:2605.31545v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve from general-purpose assistants to user-centric agents, personalization has become central to aligning model beha

safetyarxiv-cs-cl
1 Jun 2026
Model Releases

Probabilistic Precipitation Nowcasting with Rectified Flow Transformers

DGX agent

arXiv:2605.31204v1 Announce Type: new Abstract: Accurate weather forecasts are essential across various domains and are safety-critical in extreme weather conditions. Compared to simulation-based fore

model-releasesarxiv-cs-cv
1 Jun 2026
Safety

Reassessing Extractive QA Datasets at Scale: LLM-as-a-Judge and In-Depth Analyses

DGX agent

arXiv:2504.11972v3 Announce Type: replace Abstract: Extractive QA tasks are commonly evaluated using Exact Match (EM) and F1-score, but these metrics often fail to reflect true model performance. Rece

safetyarxiv-cs-cl
1 Jun 2026
Safety

Safeguarding Text-to-Image Generation via Inference-Time Prompt-Noise Optimization

DGX agent

arXiv:2412.03876v2 Announce Type: replace Abstract: Text-to-Image (T2I) diffusion models are widely recognized for their ability to generate high-quality and diverse images based on text prompts. Howe

safetyarxiv-cs-cv
1 Jun 2026
Research

SCOPE: Self-Play via Co-Evolving Policies for Open-Ended Tasks

DGX agent

arXiv:2605.31433v1 Announce Type: new Abstract: Self-play can train language models without external supervision. However, existing methods require rule-checkable answers, leaving open-ended tasks dep

researcharxiv-cs-cl
1 Jun 2026
Applications

Spatio-temporal stochastic graph-based learning for infectious disease forecasting

DGX agent

arXiv:2605.30662v1 Announce Type: new Abstract: Spatio-temporal graph-based models have typically been used to forecast new cases of infectious diseases such as COVID-19 and chickenpox outbreaks. Howe

applicationsarxiv-cs-lg
1 Jun 2026
Model Releases

Target-Agnostic Calibration under Distribution Shift with Frequency-Aware Gradient Rectification

DGX agent

arXiv:2508.19830v2 Announce Type: replace-cross Abstract: Real-world model deployments inevitably encounter distribution shifts, rendering the confidence estimates of deep neural networks highly unrel

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

The Regularizing Power of Language-Training Deepfake Detectors

DGX agent

arXiv:2605.31192v1 Announce Type: new Abstract: Recently, thanks to the advent of Multimodal-LLMs, deepfake detectors are striving not only to be generalizable but also interpretable. We propose that

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

TRACE: Discovering Task-Specific Parameter via Adaptation-Aware Probing for Continual Fine-Tuning

DGX agent

arXiv:2605.31025v1 Announce Type: new Abstract: In real-world deployment, LLMs are often adapted continually across tasks to keep LLMs up-to-date in production, where new fine-tuning should preserve p

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories

DGX agent

arXiv:2605.31308v1 Announce Type: new Abstract: Agent benchmarks increasingly record rich interaction trajectories, yet evaluation often reduces each rollout to a pass rate or reward score. We introdu

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Translation Analytics for Freelancers II: Benchmarking Local LLMs for Confidential Translation Workflows

DGX agent

arXiv:2605.31452v1 Announce Type: new Abstract: Building on our previous work, this paper develops practical, low-barrier methods for freelance translators and smaller language service providers to ev

model-releasesarxiv-cs-cl
1 Jun 2026
Research

View Space: Learning Representation across Arbitrary Graphs

DGX agent

arXiv:2512.11561v2 Announce Type: replace Abstract: Generalizing pretrained models to unseen datasets without retraining is a central challenge toward foundation models. Achieving fully inductive infe

researcharxiv-cs-lg
1 Jun 2026
Model Releases

was running some evals this weekend and claude kept trying to get me to go to bed

DGX agent

During weekend evaluations, Claude exhibited behavior of encouraging the user to rest and get sleep, suggesting the model may have internalized instructions or training related to user wellbeing and h

model-releasesyohei-nakajima--x
1 Jun 2026
Safety

What Am I Missing? Question-Answering as Hidden State Probing

DGX agent

arXiv:2605.31561v1 Announce Type: new Abstract: Test-time reasoning has become a significant field of study since the introduction of chain-of-thought reasoning in large language models (LLMs). Howeve

safetyarxiv-cs-cl
1 Jun 2026
Research

What Does Preference Learning Recover from Pairwise Comparison Data?

DGX agent

arXiv:2602.10286v2 Announce Type: replace Abstract: Pairwise preference learning is central to machine learning, with recent applications in aligning language models with human preferences. A typical

researcharxiv-cs-lg
1 Jun 2026
Model Releases

With Nemotron & Cosmos NVIDA gonna commoditise everyone's complement

DGX agent

Emad Mostaque suggests that NVIDIA's Nemotron and Cosmos models will commoditize complementary AI technologies and services in the market. The statement implies that these NVIDIA offerings will make e

model-releasesemad-mostaque--x
1 Jun 2026
Model Releases

Five million users would agree. Resetting the limits tomorrow morning to celebrate. Time to go /fast

DGX agent

Five million users would agree. Resetting the limits tomorrow morning to celebrate. Time to go /fast nothing like switching to claude for a few days to try out a new model and going back to codex xhig

model-releasessam-altman--x
31 May 2026
Agents

the market is speaking

DGX agent

the market is speaking The latest finding in the LangSmith Signal: Open Models are having a moment. 1 in 3 AI teams ran an open-weights model in April 2026, up from 1 in 5 nine months ago. The overall

agentsharrison-chase--x
30 May 2026
Model Releases

Adapting Automotive Aerodynamics Surrogates to New Vehicle Families via Transfer Learning

DGX agent

arXiv:2605.27968v1 Announce Type: cross Abstract: Deploying Scientific Machine Learning surrogates in industrial CFD workflows requires adapting pretrained models to new vehicle families without large

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

AfriScience-MT: Towards Decolonizing Science in Africa through Text Translation

DGX agent

arXiv:2605.29741v1 Announce Type: new Abstract: The dominance of colonial languages in African education and scientific communication limits how hundreds of millions of speakers of African languages a

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security

DGX agent

arXiv:2605.29801v1 Announce Type: new Abstract: Modern open-world agents such as OpenClaw exhibit powerful cross-environment execution capabilities yet introduce broad new safety risk sources. Meanwhi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Aligned but Fragile: Enhancing LLM Safety Robustness via Zeroth-Order Optimization

DGX agent

arXiv:2605.29396v1 Announce Type: new Abstract: Safety alignment for large language models (LLMs) aims to reduce harmful or unsafe behavior while preserving general utility. However, recent findings r

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Beyond Recall: Behavioral Specification as an Interpretive Layer for AI Personalization

DGX agent

arXiv:2605.28969v1 Announce Type: cross Abstract: If an AI agent makes decisions on a person's behalf, those decisions must align with its user. We introduce representational accuracy to measure how f

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Brain-IT-VQA: From Brain Signals to Answers

DGX agent

arXiv:2605.29588v1 Announce Type: cross Abstract: Decoding visual content from fMRI signals recorded while a person views images, and specifically answering questions about the seen images, is a long-

model-releasesarxiv-cs-ai
29 May 2026
Safety

Causal Interventions on Continuous Variables: A Case Study on Verb Bias in Steering Vectors for In-Context Learning

DGX agent

arXiv:2605.29971v1 Announce Type: new Abstract: Causal interventions in language model representations have largely targeted discrete features, like grammatical number. However, language models must a

safetyarxiv-cs-cl
29 May 2026
Safety

CB-SLICE: Concept-Based Interpretable Error Slice Discovery

DGX agent

arXiv:2605.29836v1 Announce Type: cross Abstract: Despite strong average-case performance, deep learning models often exhibit systematic errors on specific population groups, known as error slices. Id

safetyarxiv-cs-ai
29 May 2026
Model Releases

Citation-Closure Retrieval and Per-Rule Attribution for Real-World Regulatory Compliance Question Answering

DGX agent

arXiv:2605.29742v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) for regulatory compliance demands rigorous traceability via comprehensive citations across multi-tiered authority

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

CodeEvolve: an open source evolutionary coding agent for algorithmic discovery and optimization

DGX agent

arXiv:2510.14150v5 Announce Type: replace Abstract: We introduce CodeEvolve, an open-source framework that couples large language models with island-based evolutionary search for end-to-end algorithmi

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…531532533534535…1380
Next →