AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

Derivation Prompting: A Logic-Based Method for Improving Retrieval-Augmented Generation

DGX agent

arXiv:2605.14053v1 Announce Type: cross Abstract: The application of Large Language Models to Question Answering has shown great promise, but important challenges such as hallucinations and erroneous

model-releasesarxiv-cs-ai
15 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict

DGX agent

arXiv:2605.14473v1 Announce Type: cross Abstract: The Context-Compliance Regime in Retrieval-Augmented Generation (RAG) occurs when retrieved context dominates the final answer even when it conflicts

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Exemplar Partitioning for Mechanistic Interpretability

DGX agent

arXiv:2605.14347v1 Announce Type: new Abstract: We introduce Exemplar Partitioning (EP), an unsupervised method for constructing interpretable feature dictionaries from large language model activation

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

FlowInOne:Unifying Multimodal Generation as Image-in, Image-out Flow Matching

DGX agent

arXiv:2604.06757v2 Announce Type: replace Abstract: Multimodal generation has long been dominated by text-driven pipelines where language dictates vision but cannot reason or create within it. We chal

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Fusion-fission forecasts when AI will shift to undesirable behavior

DGX agent

arXiv:2605.14218v1 Announce Type: new Abstract: The key problem facing ChatGPT-like AI's use across society is that its behavior can shift, unnoticed, from desirable to undesirable -- encouraging self

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

GenExam: A Multidisciplinary Text-to-Image Exam

DGX agent

arXiv:2509.14232v5 Announce Type: replace Abstract: Exams are a fundamental test of expert-level intelligence and require integrated understanding, reasoning, and generation. Existing exam-style bench

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

GPart: End-to-End Isometric Fine-Tuning via Global Parameter Partitioning

DGX agent

arXiv:2605.14841v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has become the dominant paradigm for parameter-efficient fine-tuning (PEFT) of large language models (LLMs). However, its b

model-releasesarxiv-cs-ai
15 May 2026
Safety

GradShield: Alignment Preserving Finetuning

DGX agent

arXiv:2605.14194v1 Announce Type: new Abstract: Large Language Models (LLMs) pose a significant risk of safety misalignment after finetuning, as models can be compromised by both explicitly and implic

safetyarxiv-cs-cl
15 May 2026
Model Releases

GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations

DGX agent

arXiv:2605.14498v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly serve as personal assistants and workplace collaborators, where their utility depends on memory systems t

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

IPR-1: Interactive Physical Reasoner

DGX agent

arXiv:2511.15407v3 Announce Type: replace Abstract: Humans learn by observing, interacting with environments, and internalizing physics and causality. Here, we aim to ask whether an agent can similarl

model-releasesarxiv-cs-ai
15 May 2026
Safety

Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis

DGX agent

arXiv:2605.14392v1 Announce Type: new Abstract: We pursue a vision for self-improving language models in which the model does not merely generate problems or traces to imitate, but constructs the envi

safetyarxiv-cs-ai
15 May 2026
Model Releases

LoRA in LoRA: Towards Parameter-Efficient Architecture Expansion for Continual Visual Instruction Tuning

DGX agent

arXiv:2508.06202v2 Announce Type: replace-cross Abstract: Continual Visual Instruction Tuning (CVIT) enables Multimodal Large Language Models (MLLMs) to incrementally learn new tasks over time. Howeve

model-releasesarxiv-cs-ai
15 May 2026
Agents

MALLVI: A Multi-Agent Framework for Integrated Generalized Robotics Manipulation

DGX agent

arXiv:2602.16898v5 Announce Type: replace-cross Abstract: Task planning for robotic manipulation with large language models (LLMs) is an emerging area. Prior approaches rely on specialized models, fin

agentsarxiv-cs-ai
15 May 2026
Safety

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

DGX agent

arXiv:2605.14744v1 Announce Type: cross Abstract: Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal-

safetyarxiv-cs-ai
15 May 2026
Local Ai

Mistletoe: Stealthy Acceleration-Collapse Attacks on Speculative Decoding

DGX agent

arXiv:2605.14005v1 Announce Type: new Abstract: Speculative decoding has become a widely adopted technique for accelerating large language model (LLM) inference by drafting multiple candidate tokens a

local-aiarxiv-cs-cl
15 May 2026
Model Releases

MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation

DGX agent

arXiv:2605.13857v1 Announce Type: cross Abstract: The creation of cinematic-quality animal effects necessitates the precise modeling of muscle and fur dynamics, a process that remains both labor-inten

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

MUON+: Towards More Effective Muon via One Additional Normalization Step for LLM Pre-training

DGX agent

arXiv:2602.21545v3 Announce Type: replace Abstract: Muon has recently emerged as a strong optimizer for large language model pre-training, orthogonalizing the momentum matrix via Newton--Schulz polar

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Neural Signals Generate Clinical Notes in the Wild

DGX agent

arXiv:2601.22197v3 Announce Type: replace-cross Abstract: Generating clinical reports that summarize abnormal patterns, diagnostic findings, and clinical interpretations from long-term EEG recordings

model-releasesarxiv-cs-ai
15 May 2026
Research

RoSHAP: A Distributional Framework and Robust Metric for Stable Feature Attribution

DGX agent

arXiv:2605.15154v1 Announce Type: cross Abstract: Feature attribution analysis is critical for interpreting machine learning models and supporting reliable data-driven decisions. However, feature attr

researcharxiv-cs-lg
15 May 2026
Model Releases

SceneFunRI: Reasoning the Invisible for Task-Driven Functional Object Localization

DGX agent

arXiv:2605.14704v1 Announce Type: cross Abstract: In real-world scenes, target objects may reside in regions that are not visible. While humans can often infer the locations of occluded objects from c

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

SCOOTER: A Human Evaluation Framework for Unrestricted Adversarial Examples

DGX agent

arXiv:2507.07776v3 Announce Type: replace Abstract: Unrestricted adversarial attacks aim to fool computer vision models without being constrained by ell_p-norm bounds to remain imperceptible to humans

model-releasesarxiv-cs-cv
15 May 2026
Tutorials

Support Before Frequency in Discrete Diffusion

DGX agent

arXiv:2605.13999v1 Announce Type: new Abstract: Discrete diffusion models are increasingly competitive for language modeling, yet it remains unclear how their denoising objectives organize learning. A

tutorialsarxiv-cs-lg
15 May 2026
Model Releases

SWE-Chain: Benchmarking Coding Agents on Chained Release-Level Package Upgrades

DGX agent

arXiv:2605.14415v1 Announce Type: cross Abstract: Coding agents powered by large language models are increasingly expected to perform realistic software maintenance tasks beyond isolated issue resolut

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Teaching and Evaluating LLMs to Reason About Polymer Design Related Tasks

DGX agent

arXiv:2601.16312v2 Announce Type: replace-cross Abstract: Research in AI4Science has shown promise in many science applications, including polymer design. However, current LLMs are ineffective in this

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning

DGX agent

arXiv:2605.14876v1 Announce Type: cross Abstract: Despite rapid advancements, current text-to-image (T2I) models predominantly rely on a single-step generation paradigm, which struggles with complex s

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

ViMU: Benchmarking Video Metaphorical Understanding

DGX agent

arXiv:2605.14607v1 Announce Type: new Abstract: Any new medium, once it emerges, is used for more than the transmission of overt content alone. The information it carries typically operates on two lev

model-releasesarxiv-cs-cv
15 May 2026
Safety

Vision-LLMs for Spatiotemporal Traffic Forecasting

DGX agent

arXiv:2510.11282v2 Announce Type: replace Abstract: Accurate spatiotemporal traffic forecasting is a critical prerequisite for proactive resource management in dense urban mobile networks. While large

safetyarxiv-cs-lg
15 May 2026
Model Releases

When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution

DGX agent

arXiv:2605.14504v1 Announce Type: new Abstract: Long-horizon household tasks demand robust high-level planning and sustained reasoning capabilities, which are largely overlooked by existing embodied A

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

A3 : an Analytical Low-Rank Approximation Framework for Attention

DGX agent

arXiv:2505.12942v4 Announce Type: replace-cross Abstract: Large language models have demonstrated remarkable performance; however, their massive parameter counts make deployment highly expensive. Low-

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Building Interactive Real-Time Agents with Asynchronous I/O and Speculative Tool Calling

DGX agent

arXiv:2605.13360v1 Announce Type: new Abstract: There is a growing demand for agentic AI technologies for a range of downstream applications like customer service and personal assistants. For applicat

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Children's English Reading Story Generation via Supervised Fine-Tuning of Compact LLMs with Controllable Difficulty and Safety

DGX agent

arXiv:2605.13709v1 Announce Type: cross Abstract: Large Language Models (LLMs) are widely applied in educational practices, such as for generating children's stories. However, the generated stories ar

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Collaborative Parameter Learning: Mitigating Forgetting via Parameter-Level Gradient Analysis

DGX agent

arXiv:2601.21577v2 Announce Type: replace Abstract: Catastrophic forgetting during knowledge injection impairs the ability of large language models to acquire new knowledge without overwriting previou

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Dense vs Sparse Pretraining at Tiny Scale: Active-Parameter vs Total-Parameter Matching

DGX agent

arXiv:2605.13769v1 Announce Type: cross Abstract: We study dense and mixture-of-experts (MoE) transformers in a tiny-scale pretraining regime under a shared LLaMA-style decoder training recipe. The sp

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Descriptive Collision in Sparse Autoencoder Auto-Interpretability: When One Explanation Describes Many Features

DGX agent

arXiv:2605.12874v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) are now standard tools for decomposing language model activations into interpretable features, and automated interpretability

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack

DGX agent

arXiv:2605.12673v1 Announce Type: new Abstract: Agent benchmarks have become the de facto measure of frontier AI competence, guiding model selection, investment, and deployment. However, reward hackin

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion

DGX agent

arXiv:2605.11679v2 Announce Type: replace Abstract: In the realm of multi-objective alignment for large language models, balancing disparate human preferences often manifests as a zero-sum conflict. S

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Exploring Multimodal LMMs for Online Episodic Memory Question Answering on the Edge

DGX agent

arXiv:2602.22455v2 Announce Type: replace Abstract: We investigate the feasibility of using Multimodal Large Language Models (MLLMs) for real-time online episodic memory question answering. While clou

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

FOAM: Blocked State Folding for Memory-Efficient LLM Training

DGX agent

arXiv:2512.07112v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated remarkable performance due to their large parameter counts and extensive training data. However

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

HCSG: Human-Centric Semantic-Geometric Reasoning for Vision-Language Navigation

DGX agent

arXiv:2605.13321v1 Announce Type: new Abstract: VLN has achieved remarkable progress by scaling data and model capacity. However, the assumption of a static environment breaks down in real-world indoo

model-releasesarxiv-cs-ro
14 May 2026
Model Releases

LIFT: Last-Mile Fine-Tuning for Table Explicitation

DGX agent

arXiv:2605.13424v1 Announce Type: new Abstract: We propose last-mile fine-tuning, or Lift, a pipeline in which a pre-trained large language model extracts an initial table from unstructured clipboard

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

MedCore: Boundary-Preserving Medical Core Pruning for MedSAM

DGX agent

arXiv:2605.13688v1 Announce Type: new Abstract: Medical segmentation foundation models such as SAM and MedSAM provide strong prompt-driven segmentation, but their image encoders are still too large fo

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence

DGX agent

arXiv:2605.12703v1 Announce Type: cross Abstract: We introduce MMCL-Bench, a benchmark for multimodal context learning: learning task-local rules, procedures, and empirical patterns from visual or mix

model-releasesarxiv-cs-ai
14 May 2026
Safety

Pareto-Guided Optimal Transport for Multi-Reward Alignment

DGX agent

arXiv:2605.13155v1 Announce Type: new Abstract: Text-to-image generation models have achieved remarkable progress in preference optimization, yet achieving robust alignment across diverse reward model

safetyarxiv-cs-cv
14 May 2026
Model Releases

Phasor Memory Networks: Stable Backpropagation Through Time for Scalable Explicit Memory

DGX agent

arXiv:2605.13370v1 Announce Type: new Abstract: For over a decade, explicit memory architectures like the Neural Turing Machine have remained theoretically appealing yet practically intractable for la

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Pitfalls of Unlabeled Disagreement-Based Drift Detection in Streaming Tree Ensembles

DGX agent

arXiv:2605.12803v1 Announce Type: new Abstract: Detecting concept drift in high-speed data streams remains challenging, particularly when models must operate on unlabeled data and avoid false alarms c

model-releasesarxiv-cs-lg
14 May 2026
Safety

Quantifying LLM Safety Degradation Under Repeated Attacks Using Survival Analysis

DGX agent

arXiv:2605.12869v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed in a wide range of applications, yet remain vulnerable to adversarial jailbreak attacks that ci

safetyarxiv-cs-ai
14 May 2026
Model Releases

REAP the Experts: Why Pruning Prevails for One-Shot MoE compression

DGX agent

arXiv:2510.13999v3 Announce Type: replace-cross Abstract: Sparsely-activated Mixture-of-Experts (SMoE) models offer efficient pre-training and low latency but their large parameter counts create signi

model-releasesarxiv-cs-ai
14 May 2026
Safety

Revealing the Gap in Human and VLM Scene Perception through Counterfactual Semantic Saliency

DGX agent

arXiv:2605.13047v1 Announce Type: cross Abstract: Evaluating whether large vision-language models (VLMs) align with human perception for high-level semantic scene comprehension remains a challenge. Tr

safetyarxiv-cs-ai
14 May 2026
← Previous
1…369370371372373…1074
Next →