AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

TRACE: Topology-aware Reconstruction of Accidents in CARLA for AV Evaluation

DGX agent

arXiv:2604.22068v1 Announce Type: cross Abstract: Validating Autonomous Vehicles (AVs) requires exposure to rare, safety-critical scenarios, infrequent in routine driving data. Existing benchmarks add

model-releasesarxiv-cs-ro
27 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TreeCoder: Systematic Exploration and Optimisation of Decoding and Constraints for LLM Code Generation

DGX agent

arXiv:2511.22277v2 Announce Type: replace Abstract: Large language models (LLMs) have shown remarkable ability to generate code, yet their outputs often violate syntactic or semantic constraints when

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

TS-Arena -- A Live Forecast Pre-Registration Platform

DGX agent

arXiv:2512.20761v3 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) are transforming the field of forecasting. However, evaluating them on historical data is increasingly d

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents

DGX agent

arXiv:2602.07038v2 Announce Type: replace-cross Abstract: Key Information Extraction (KIE) from real-world documents remains challenging due to substantial variations in layout structures, visual qual

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Universal Transformers Need Memory: Depth-State Trade-offs in Adaptive Recursive Reasoning

DGX agent

arXiv:2604.21999v1 Announce Type: cross Abstract: We study learned memory tokens as computational scratchpad for a single-block Universal Transformer (UT) with Adaptive Computation Time (ACT) on Sudok

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

UR^2: Unify RAG and Reasoning through Reinforcement Learning

DGX agent

arXiv:2508.06165v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown strong capabilities through two complementary paradigms: Retrieval-Augmented Generation (RAG) for know

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

When AI Speaks, Whose Values Does It Express? A Cross-Cultural Audit of Individualism-Collectivism Bias in Large Language Models

DGX agent

arXiv:2604.22153v1 Announce Type: cross Abstract: When you ask an AI assistant for advice about your career, your marriage, or a conflict with your family, does it give you the same answer regardless

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

When Cow Urine Cures Constipation on YouTube: Limits of LLMs in Detecting Culture-specific Health Misinformation

DGX agent

arXiv:2604.22002v1 Announce Type: new Abstract: Social media platforms have become primary channels for health information in the Global South. Using gomutra (cow urine) discourse on YouTube in India

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

When Does LLM Self-Correction Help? A Control-Theoretic Markov Diagnostic and Verify-First Intervention

DGX agent

arXiv:2604.22273v1 Announce Type: new Abstract: Iterative self-correction is widely used in agentic LLM systems, but when repeated refinement helps versus hurts remains unclear. We frame self-correcti

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Wiggle and Go! System Identification for Zero-Shot Dynamic Rope Manipulation

DGX agent

arXiv:2604.22102v1 Announce Type: cross Abstract: Many robotic tasks are unforgiving; a single mistake in a dynamic throw can lead to unacceptable delays or unrecoverable failure. To mitigate this, we

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

A Dynamic Framework for Grid Adaptation in Kolmogorov-Arnold Networks

DGX agent

arXiv:2601.18672v3 Announce Type: replace Abstract: Kolmogorov-Arnold Networks (KANs) have recently demonstrated promising potential in scientific machine learning, partly due to their capacity for gr

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

A Green-Integral-Constrained Neural Solver with Stochastic Physics-Informed Regularization

DGX agent

arXiv:2604.21411v1 Announce Type: new Abstract: Standard physics-informed neural networks (PINNs) struggle to simulate highly oscillatory Helmholtz solutions in heterogeneous media because pointwise m

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

A-IC3: Learning-Guided Adaptive Inductive Generalization for Hardware Model Checking

DGX agent

arXiv:2604.21688v1 Announce Type: cross Abstract: The IC3 algorithm represents the state-of-the-art (SOTA) hardware model checking technique, owing to its robust performance and scalability. A signifi

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair

DGX agent

arXiv:2604.21579v1 Announce Type: cross Abstract: LLM-based automated program repair (APR) techniques have shown promising results in reducing debugging costs. However, prior results can be affected b

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

A-THENA: Early Intrusion Detection for IoT with Time-Aware Hybrid Encoding and Network-Specific Augmentation

DGX agent

arXiv:2604.21623v1 Announce Type: cross Abstract: The proliferation of Internet of Things (IoT) devices has significantly expanded attack surfaces, making IoT ecosystems particularly susceptible to so

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Absorber LLM: Harnessing Causal Synchronization for Test-Time Training

DGX agent

arXiv:2604.20915v1 Announce Type: cross Abstract: Transformers suffer from a high computational cost that grows with sequence length for self-attention, making inference in long streams prohibited by

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Adaptive Defense Orchestration for RAG: A Sentinel-Strategist Architecture against Multi-Vector Attacks

DGX agent

arXiv:2604.20932v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are increasingly deployed in sensitive domains such as healthcare and law, where they rely on private, do

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

ADS-POI: Agentic Spatiotemporal State Decomposition for Next Point-of-Interest Recommendation

DGX agent

arXiv:2604.20846v1 Announce Type: cross Abstract: Next point-of-interest (POI) recommendation requires modeling user mobility as a spatiotemporal sequence, where different behavioral factors may evolv

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

AEL: Agent Evolving Learning for Open-Ended Environments

DGX agent

arXiv:2604.21725v1 Announce Type: cross Abstract: LLM agents increasingly operate in open-ended environments spanning hundreds of sequential episodes, yet they remain largely stateless: each task is s

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

AFRILANGTUTOR: Advancing Language Tutoring and Culture Education in Low-Resource Languages with Large Language Models

DGX agent

arXiv:2604.20996v1 Announce Type: new Abstract: How can language learning systems be developed for languages that lack sufficient training resources? This challenge is increasingly faced by developers

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security

DGX agent

arXiv:2601.18491v2 Announce Type: replace Abstract: The rise of AI agents introduces complex safety and security challenges arising from autonomous tool use and environmental interactions. Current gua

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

AITP: Traffic Accident Responsibility Allocation via Multimodal Large Language Models

DGX agent

arXiv:2604.20878v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress in Traffic Accident Detection (TAD) and Traffic Accident Understanding (TAU).

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

APCoTTA: Continual Test-Time Adaptation for Semantic Segmentation of Airborne LiDAR Point Clouds

DGX agent

arXiv:2505.09971v3 Announce Type: replace Abstract: Airborne laser scanning (ALS) point cloud semantic segmentation is a fundamental task for large-scale 3D scene understanding. Fixed models deployed

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

ARFBench: Benchmarking Time Series Question Answering Ability for Software Incident Response

DGX agent

arXiv:2604.21199v1 Announce Type: cross Abstract: Time series question-answering (TSQA), in which we ask natural language questions to infer and reason about properties of time series, is a promising

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

AUDITA: A New Dataset to Audit Humans vs. AI Skill at Audio QA

DGX agent

arXiv:2604.21766v1 Announce Type: new Abstract: Existing audio question answering benchmarks largely emphasize sound event classification or caption-grounded queries, often enabling models to succeed

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Beyond Accuracy: A Stability-Aware Metric for Multi-Horizon Forecasting

DGX agent

arXiv:2601.10863v3 Announce Type: replace Abstract: Traditional time series forecasting methods optimize for accuracy alone. This objective neglects temporal consistency, in other words, how consisten

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Beyond N-gram: Data-Aware X-GRAM Extraction for Efficient Embedding Parameter Scaling

DGX agent

arXiv:2604.21724v1 Announce Type: new Abstract: Large token-indexed lookup tables provide a compute-decoupled scaling path, but their practical gains are often limited by poor parameter efficiency and

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Beyond Pixels: Introspective and Interactive Grounding for Visualization Agents

DGX agent

arXiv:2604.21134v1 Announce Type: new Abstract: Vision-Language Models (VLMs) frequently misread values, hallucinate details, and confuse overlapping elements in charts. Current approaches rely solely

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Beyond Single Plots: A Benchmark for Question Answering on Multi-Charts

DGX agent

arXiv:2604.21344v1 Announce Type: cross Abstract: Charts are widely used to present complex information. Deriving meaningful insights in real-world contexts often requires interpreting multiple relate

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

BioMiner: A Multi-modal System for Automated Mining of Protein-Ligand Bioactivity Data from Literature

DGX agent

arXiv:2604.21508v1 Announce Type: new Abstract: Protein-ligand bioactivity data published in the literature are essential for drug discovery, yet manual curation struggles to keep pace with rapidly gr

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs

DGX agent

arXiv:2604.20945v1 Announce Type: cross Abstract: Effective safety auditing of large language models (LLMs) demands tools that go beyond black-box probing and systematically uncover vulnerabilities ro

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Building a Precise Video Language with Human-AI Oversight

DGX agent

arXiv:2604.21718v1 Announce Type: cross Abstract: Video-language models (VLMs) learn to reason about the dynamic visual world through natural language. We introduce a suite of open datasets, benchmark

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Can MLLMs 'Read' What is Missing?

DGX agent

arXiv:2604.21277v1 Announce Type: new Abstract: We introduce MMTR-Bench, a benchmark designed to evaluate the intrinsic ability of Multimodal Large Language Models (MLLMs) to reconstruct masked text d

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

CAP: Controllable Alignment Prompting for Unlearning in LLMs

DGX agent

arXiv:2604.21251v1 Announce Type: cross Abstract: Large language models (LLMs) trained on unfiltered corpora inherently risk retaining sensitive information, necessitating selective knowledge unlearni

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

CaST-POI: Candidate-Conditioned Spatiotemporal Modeling for Next POI Recommendation

DGX agent

arXiv:2604.20845v1 Announce Type: cross Abstract: Next Point-of-Interest (POI) recommendation plays a crucial role in location-based services by predicting users' future mobility patterns. Existing me

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents

DGX agent

arXiv:2604.21308v1 Announce Type: cross Abstract: Enterprise LLM agents can dramatically improve workplace productivity, but their core capability, retrieving and using internal context to act on a us

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

CLT-Optimal Parameter Error Bounds for Linear System Identification

DGX agent

arXiv:2604.21270v1 Announce Type: cross Abstract: There has been remarkable progress over the past decade in establishing finite-sample, non-asymptotic bounds on recovering unknown system parameters f

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Cognitive Amplification vs Cognitive Delegation in Human-AI Systems: A Metric Framework

DGX agent

arXiv:2603.18677v2 Announce Type: replace-cross Abstract: Artificial intelligence is increasingly embedded in human decision making. In some cases, it enhances human reasoning. In others, it fosters e

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Compliance Moral Hazard and the Backfiring Mandate

DGX agent

arXiv:2604.21789v1 Announce Type: cross Abstract: Competing firms that serve shared customer populations face a fundamental information aggregation problem: each firm holds fragmented signals about ri

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Concurrence: A dependence criterion for time series, applied to biological data

DGX agent

arXiv:2512.16001v2 Announce Type: replace-cross Abstract: Measuring the statistical dependence between observed signals is a primary tool for scientific discovery. However, biological systems often ex

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Context Is What You Need: The Maximum Effective Context Window for Real World Limits of LLMs

DGX agent

arXiv:2509.21361v2 Announce Type: replace-cross Abstract: Large language model (LLM) providers boast big numbers for maximum context window sizes. To test the real world use of context windows, we 1)

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors

DGX agent

arXiv:2604.21241v1 Announce Type: cross Abstract: Vision--Language--Action (VLA) models often use intermediate representations to connect multimodal inputs with continuous control, yet spatial guidanc

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Counterfactual Segmentation Reasoning: Diagnosing and Mitigating Pixel-Grounding Hallucination

DGX agent

arXiv:2506.21546v4 Announce Type: replace-cross Abstract: Segmentation Vision-Language Models (VLMs) have significantly advanced grounded visual understanding, yet they remain prone to pixel-grounding

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Cross-Entropy Is Load-Bearing: A Pre-Registered Scope Test of the K-Way Energy Probe on Bidirectional Predictive Coding

DGX agent

arXiv:2604.21286v1 Announce Type: cross Abstract: Cacioli (2026) showed that the K-way energy probe on standard discriminative predictive coding networks reduces approximately to a monotone function o

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Cross-Session Threats in AI Agents: Benchmark, Evaluation, and Algorithms

DGX agent

arXiv:2604.21131v1 Announce Type: cross Abstract: AI-agent guardrails are memoryless: each message is judged in isolation, so an adversary who spreads a single attack across dozens of sessions slips p

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

CSC: Turning the Adversary's Poison against Itself

DGX agent

arXiv:2604.21416v1 Announce Type: cross Abstract: Poisoning-based backdoor attacks pose significant threats to deep neural networks by embedding triggers in training data, causing models to misclassif

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Data-Driven Open-Loop Simulation for Digital-Twin Operator Decision Support in Wastewater Treatment

DGX agent

arXiv:2604.20935v1 Announce Type: cross Abstract: Wastewater treatment plants (WWTPs) need digital-twin-style decision support tools that can simulate plant response under prescribed control plans, to

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

DAVIS: OOD Detection via Dominant Activations and Variance for Increased Separation

DGX agent

arXiv:2601.22703v2 Announce Type: replace Abstract: Detecting out-of-distribution (OOD) inputs is a critical safeguard for deploying machine learning models in the real world. However, most post-hoc d

model-releasesarxiv-cs-cv
24 Apr 2026
← Previous
1…300301302303304…357
Next →