AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Safety

BARRED: Synthetic Training of Custom Policy Guardrails via Asymmetric Debate

DGX agent

arXiv:2604.25203v1 Announce Type: new Abstract: Deploying guardrails for custom policies remains challenging, as generic safety models fail to capture task-specific requirements, while prompting LLMs

safetyarxiv-cs-cl
29 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

BLASST: Dynamic BLocked Attention Sparsity via Softmax Thresholding

DGX agent

arXiv:2512.12087v3 Announce Type: replace Abstract: The growing demand for long-context inference capabilities in Large Language Models (LLMs) has intensified the computational and memory bottlenecks

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Conditional Flow Matching for Probabilistic Downscaling of Maximum 3-day Snowfall in Alaska

DGX agent

arXiv:2604.25172v1 Announce Type: cross Abstract: Precipitation in complex terrain is governed by orographic processes operating at scales of a few kilometers, yet climate models typically run at reso

researcharxiv-cs-lg
29 Apr 2026
Tutorials

DDA-Thinker: Decoupled Dual-Atomic Reinforcement Learning for Reasoning-Driven Image Editing

DGX agent

arXiv:2604.25477v1 Announce Type: new Abstract: Recent image editing models have achieved strong visual fidelity but often struggle with tasks requiring complex reasoning. To investigate and enhance t

tutorialsarxiv-cs-cv
29 Apr 2026
Research

DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference

DGX agent

arXiv:2510.19669v4 Announce Type: replace Abstract: Recent reasoning Large Language Models (LLMs) demonstrate remarkable problem-solving abilities but often generate long thinking traces whose utility

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Doing More With Less: Revisiting the Effectiveness of LLM Pruning for Test-Time Scaling

DGX agent

arXiv:2604.25098v1 Announce Type: cross Abstract: While current Large Language Models (LLMs) exhibit remarkable reasoning capabilities through test-time compute scaling (TTS), their massive parameter

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

FED-FSTQ: Fisher-Guided Token Quantization for Communication-Efficient Federated Fine-Tuning of LLMs on Edge Devices

DGX agent

arXiv:2604.25421v1 Announce Type: new Abstract: Federated fine-tuning provides a practical route to adapt large language models (LLMs) on edge devices without centralizing private data, yet in mobile

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Golden RPG: Confidence-Adaptive Region-Aware Noise for Compositional Text-to-Image Generation

DGX agent

arXiv:2604.25314v1 Announce Type: new Abstract: Compositional text-to-image (T2I) generation requires a model to honour multiple sub-prompts that describe distinct image regions. Recent work shows tha

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Improving LLM Predictions via Inter-Layer Structural Encoders

DGX agent

arXiv:2603.22665v2 Announce Type: replace Abstract: The standard practice in Large Language Models (LLMs) is to base predictions on final-layer representations. However, intermediate layers encode com

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Limited Linguistic Diversity in Embodied AI Datasets

DGX agent

arXiv:2601.03136v2 Announce Type: replace Abstract: Language plays a critical role in Vision-Language-Action (VLA) models, yet the linguistic characteristics of the datasets used to train and evaluate

researcharxiv-cs-cl
29 Apr 2026
Model Releases

LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation

DGX agent

arXiv:2604.25665v1 Announce Type: new Abstract: Reliable evaluation of large language model (LLM)-generated summaries remains an open challenge, particularly across heterogeneous domains and document

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition

DGX agent

arXiv:2512.07348v2 Announce Type: replace Abstract: In controllable image generation, synthesizing coherent and consistent images from multiple reference inputs, i.e., Multi-Image Composition (MICo),

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Quantifying and Mitigating Socially Desirable Responding in LLMs: A Desirability-Matched Graded Forced-Choice Psychometric Study

DGX agent

arXiv:2602.17262v2 Announce Type: replace Abstract: Human self-report questionnaires are increasingly used in NLP to benchmark and audit large language models (LLMs), from persona consistency to safet

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

RCProb: Probabilistic Rule Extraction for Efficient Simplification of Tree Ensembles

DGX agent

arXiv:2604.25304v1 Announce Type: new Abstract: Tree ensembles are widely used in industrial machine learning due to their strong predictive performance and efficient training procedures. However, as

model-releasesarxiv-cs-lg
29 Apr 2026
Applications

Relational In-Context Learning via Synthetic Pre-training with Structural Prior

DGX agent

arXiv:2603.03805v2 Announce Type: replace Abstract: Relational Databases (RDBs) are the backbone of modern business, yet they lack foundation models comparable to those in text or vision. A key obstac

applicationsarxiv-cs-lg
29 Apr 2026
Safety

ReSim: Reliable World Simulation for Autonomous Driving

DGX agent

arXiv:2506.09981v2 Announce Type: replace Abstract: How can we reliably simulate future driving scenarios under a wide range of ego driving behaviors? Recent driving world models, developed exclusivel

safetyarxiv-cs-cv
29 Apr 2026
Research

Sensitivity-Based Tube NMPC for Cooperative Aerial Structures Under Parametric Uncertainty

DGX agent

arXiv:2604.25766v1 Announce Type: new Abstract: This paper presents a sensitivity-based tube Nonlinear Model Predictive Control (NMPC) framework for cooperative aerial chains under bounded parametric

researcharxiv-cs-ro
29 Apr 2026
Research

TopoMamba: Topology-Aware Scanning and Fusion for Segmenting Heterogeneous Medical Visual Media

DGX agent

arXiv:2604.25545v1 Announce Type: new Abstract: Visual state-space models (SSMs) have shown strong potential for medical image segmentation, yet their effectiveness is often limited by two practical i

researcharxiv-cs-cv
29 Apr 2026
Model Releases

When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs

DGX agent

arXiv:2510.07499v2 Announce Type: replace Abstract: Recent Long-Context Language Models (LCLMs) can process hundreds of thousands of tokens in a single prompt, enabling new opportunities for knowledge

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Aligning with Your Own Voice: Self-Corrected Preference Learning for Hallucination Mitigation in LVLMs

DGX agent

arXiv:2604.24395v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) frequently suffer from hallucinations. Existing preference learning-based approaches largely rely on proprietary mo

safetyarxiv-cs-ai
28 Apr 2026
Safety

Analytica: Soft Propositional Reasoning for Robust and Scalable LLM-Driven Analysis

DGX agent

arXiv:2604.23072v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly tasked with complex real-world analysis (e.g., in financial forecasting, scientific discovery), yet t

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

AsyncShield: A Plug-and-Play Edge Adapter for Asynchronous Cloud-based VLA Navigation

DGX agent

arXiv:2604.24086v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models have been demonstrated possessing strong zero-shot generalization for robot control, their massive parameter

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation

DGX agent

arXiv:2604.24665v1 Announce Type: cross Abstract: This paper investigates whether source trustworthiness shapes Turkish evidential morphology and whether large language models (LLMs) track this sensit

local-aiarxiv-cs-ai
28 Apr 2026
Model Releases

Beyond Local vs. External: A Game-Theoretic Framework for Trustworthy Knowledge Acquisition

DGX agent

arXiv:2604.23413v1 Announce Type: new Abstract: Cloud-hosted Large Language Models (LLMs) offer unmatched reasoning capabilities and dynamic knowledge, yet submitting raw queries to these external ser

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

BIR-Adapter: A parameter-efficient diffusion adapter for blind image restoration

DGX agent

arXiv:2509.06904v3 Announce Type: replace Abstract: We introduce the BIR-Adapter, a parameter-efficient diffusion adapter for blind image restoration. Diffusion-based restoration methods have demonstr

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents

DGX agent

arXiv:2604.23781v1 Announce Type: new Abstract: Language-model agents are increasingly used as persistent coworkers that assist users across multiple working days. During such workflows, the surroundi

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

Computational Design and Co-Robotic Fabrication for Material Reuse in Architecture

DGX agent

arXiv:2604.24648v1 Announce Type: new Abstract: Climate change and resource depletion demand a shift from the dominant linear 'take-make-use-dispose' paradigm of construction toward circular, low-wast

applicationsarxiv-cs-ro
28 Apr 2026
Research

Continual Calibration: Coverage Can Collapse Before Accuracy in Lifelong LLM Fine-Tuning

DGX agent

arXiv:2604.23987v1 Announce Type: new Abstract: Continual learning for large language models is typically evaluated through accuracy retention under sequential fine-tuning. We argue that this perspect

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Cortex-Inspired Continual Learning: Unsupervised Instantiation and Recovery of Functional Task Networks

DGX agent

arXiv:2604.24637v1 Announce Type: cross Abstract: Block-sequential continual learning demands that a single model both protect prior solutions from catastrophic forgetting and efficiently infer at inf

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Coverage-Based Calibration for Post-Training Quantization via Weighted Set Cover over Outlier Channels

DGX agent

arXiv:2604.24008v1 Announce Type: new Abstract: Post-Training Quantization (PTQ) compresses large language models to low bit-widths using a small calibration set, and its quality depends strongly on w

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

DecompKAN: Decomposed Patch-KAN for Long-Term Time Series Forecasting

DGX agent

arXiv:2604.23968v1 Announce Type: cross Abstract: Accurate time series forecasting in scientific domains such as climate modeling, physiological monitoring, and energy systems benefits from both compe

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

Deep Learning of Solver-Aware Turbulence Closures from Nudged LES Dynamics

DGX agent

arXiv:2604.23874v1 Announce Type: cross Abstract: Deep learning approaches have shown remarkable promise in turbulence closure modeling for large eddy simulations (LES). The differentiable physics par

tutorialsarxiv-cs-lg
28 Apr 2026
Model Releases

Defusing the Trigger: Plug-and-Play Defense for Backdoored LLMs via Tail-Risk Intrinsic Geometric Smoothing

DGX agent

arXiv:2604.24162v1 Announce Type: cross Abstract: Defending against backdoor attacks in large language models remains a critical practical challenge. Existing defenses mitigate these threats but typic

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Deploy DINO with Many-to-Many Association

DGX agent

arXiv:2604.23670v1 Announce Type: new Abstract: Motivated by the limited generalization of supervised image matching models to unseen image domains, we explore the zero-shot deployment of DINO feature

model-releasesarxiv-cs-cv
28 Apr 2026
Research

DiffuSAM: Diffusion-Based Prompt-Free SAM2 for Few-Shot and Source-Free Medical Image Segmentation

DGX agent

arXiv:2604.24719v1 Announce Type: new Abstract: Segmentation models such as Segment Anything Model (SAM) and SAM2 achieve strong prompt-driven zero-shot performance. However, their training on natural

researcharxiv-cs-cv
28 Apr 2026
Model Releases

EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs

DGX agent

arXiv:2604.23348v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and generation, and are increasingly used in

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Expert Evaluation of LLM's Open-Ended Legal Reasoning on the Japanese Bar Exam Writing Task

DGX agent

arXiv:2604.23730v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong performance on legal benchmarks, including multiple-choice components of bar exams. However, their capaci

applicationsarxiv-cs-ai
28 Apr 2026
Safety

Explanation Quality Assessment as Ranking with Listwise Rewards

DGX agent

arXiv:2604.24176v1 Announce Type: new Abstract: We reformulate explanation quality assessment as a ranking problem rather than a generation problem. Instead of optimizing models to produce a single 'b

safetyarxiv-cs-ai
28 Apr 2026
Research

FlashOverlap: Minimizing Tail Latency in Communication Overlap for Distributed LLM Training

DGX agent

arXiv:2604.24013v1 Announce Type: cross Abstract: The rapid growth in the size of large language models has necessitated the partitioning of computational workloads across accelerators such as GPUs, T

researcharxiv-cs-cv
28 Apr 2026
Model Releases

Generating Place-Based Compromises Between Two Points of View

DGX agent

arXiv:2604.24536v1 Announce Type: new Abstract: Large Language Models (LLMs) excel academically but struggle with social intelligence tasks, such as creating good compromises. In this paper, we presen

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Green Prompting: Characterizing Prompt-driven Energy Costs of LLM Inference

DGX agent

arXiv:2503.10666v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become widely used across various domains spanning search engines, code generation, and text creation. Howev

researcharxiv-cs-ai
28 Apr 2026
Safety

Isotonic Layer: A Unified Framework for Recommendation Calibration and Debiasing

DGX agent

arXiv:2603.06589v2 Announce Type: replace-cross Abstract: Model calibration and debiasing are fundamental yet operationally expensive challenges in large-scale recommendation systems. Existing approac

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

KLong: Training LLM Agent for Extremely Long-horizon Tasks

DGX agent

arXiv:2602.17547v3 Announce Type: replace Abstract: This paper introduces KLong, an open-source LLM agent trained to solve extremely long-horizon tasks. The principle is to first cold-start the model

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MEG-RAG: Quantifying Multi-modal Evidence Grounding for Evidence Selection in RAG

DGX agent

arXiv:2604.24564v1 Announce Type: new Abstract: Multimodal Retrieval-Augmented Generation (MRAG) addresses key limitations of Multimodal Large Language Models (MLLMs), such as hallucination and outdat

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

MEMCoder: Multi-dimensional Evolving Memory for Private-Library-Oriented Code Generation

DGX agent

arXiv:2604.24222v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at general code generation, but their performance drops sharply in enterprise settings that rely on internal privat

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MetaErr: Towards Predicting Error Patterns in Deep Neural Networks

DGX agent

arXiv:2604.23289v1 Announce Type: cross Abstract: Due to the unprecedented success of deep learning, it has become an integral component in several multimedia computing applications in todays world. U

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Mobile-R1: Towards Interactive Capability for VLM-Based Mobile Agent via Systematic Training

DGX agent

arXiv:2506.20332v4 Announce Type: replace Abstract: Vision-language model-based mobile agents have gained the ability to understand complex instructions and mobile screenshots, benefiting from reinfor

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation

DGX agent

arXiv:2604.23789v1 Announce Type: new Abstract: While video foundation models excel at single-shot generation, real-world cinematic storytelling inherently relies on complex multi-shot sequencing. Fur

model-releasesarxiv-cs-cv
28 Apr 2026
← Previous
1…376377378379380…1074
Next →