AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,952 results
14 Apr 2026

A-IO: Adaptive Inference Orchestration for Memory-Bound NPUs

ResearchDGX agent

arXiv:2604.09752v1 Announce Type: cross Abstract: During the deployment of Large Language Models (LLMs), the autoregressive decoding phase on heterogeneous NPU platforms (e.g., Ascend 910B) faces seve

AI-enhanced tuning of quantum dot Hamiltonians toward Majorana modes

Model ReleasesDGX agent

arXiv:2601.02149v3 Announce Type: replace-cross Abstract: We propose a neural network-based model capable of learning the broad landscape of working regimes in quantum dot simulators, and using this k

AIM-Bench: Benchmarking and Improving Affective Image Manipulation via Fine-Grained Hierarchical Control

Model ReleasesDGX agent

arXiv:2604.10454v1 Announce Type: new Abstract: Affective Image Manipulation (AIM) aims to evoke specific emotions through targeted editing. Current image editing benchmarks primarily focus on object-

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AIMER: Calibration-Free Task-Agnostic MoE Pruning

Model ReleasesDGX agent

arXiv:2603.18492v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) language models increase parameter capacity without proportional per-token compute, but the deployment still requires stori

BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs

Model ReleasesDGX agent

arXiv:2604.10528v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) demonstrate remarkable zero-shot recognition capabilities across a diverse spectrum of multimodal tasks, it yet rema

Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs

SafetyDGX agent

arXiv:2604.10585v1 Announce Type: cross Abstract: Modern large language models (LLMs) are increasingly fine-tuned via reinforcement learning from human feedback (RLHF) or related reward optimisation s

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

Model ReleasesDGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection

Local AiDGX agent

arXiv:2604.11790v1 Announce Type: cross Abstract: Tool-augmented Large Language Model (LLM) agents have demonstrated impressive capabilities in automating complex, multi-step real-world tasks, yet rem

CodaRAG: Connecting the Dots with Associativity Inspired by Complementary Learning

ResearchDGX agent

arXiv:2604.10426v1 Announce Type: cross Abstract: Large Language Models (LLMs) struggle with knowledge-intensive tasks due to hallucinations and fragmented reasoning over dispersed information. While

CodeQuant: Unified Clustering and Quantization for Enhanced Outlier Smoothing in Low-Precision Mixture-of-Experts

HardwareDGX agent

arXiv:2604.10496v1 Announce Type: new Abstract: Outliers have emerged as a fundamental bottleneck in preserving accuracy for low-precision large models, particularly within Mixture-of-Experts (MoE) ar

Correct me if I’m wrong: Ollama can’t fine tune like Unsloth Studio

Local AiDGX agent

Ollama is a local inference engine designed for running pre-built LLMs on your own machine, and it does not include fine-tuning capabilities — this distinction is correct. Unsloth (and its Unsloth Stu

CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing

Model ReleasesDGX agent

arXiv:2506.18438v2 Announce Type: replace Abstract: Editing natural images using textual descriptions in text-to-image diffusion models remains a significant challenge, particularly in achieving consi

CricBench: A Multilingual Benchmark for Evaluating LLMs in Cricket Analytics

Model ReleasesDGX agent

arXiv:2512.21877v3 Announce Type: replace-cross Abstract: Cricket is the second most popular sport worldwide, with billions of fans seeking advanced statistical insights unavailable through standard w

DecepGPT: Schema-Driven Deception Detection with Multicultural Datasets and Robust Multimodal Learning

Model ReleasesDGX agent

arXiv:2603.23916v2 Announce Type: replace-cross Abstract: Multimodal deception detection aims to identify deceptive behavior by analyzing audiovisual cues for forensics and security. In these high-sta

Defending against Backdoor Attacks via Module Switching

ResearchDGX agent

arXiv:2504.05902v2 Announce Type: replace-cross Abstract: Backdoor attacks pose a serious threat to deep neural networks (DNNs), allowing adversaries to implant triggers for hidden behaviors in infere

Delta Rectified Flow Sampling for Text-to-Image Editing

Model ReleasesDGX agent

arXiv:2509.05342v3 Announce Type: replace Abstract: We propose Delta Rectified Flow Sampling (DRFS), a novel inversion-free, path-aware editing framework within rectified flow models for text-to-image

DERM-3R: A Resource-Efficient Multimodal Agents Framework for Dermatologic Diagnosis and Treatment in Real-World Clinical Settings

Model ReleasesDGX agent

arXiv:2604.09596v1 Announce Type: new Abstract: Dermatologic diseases impose a large and growing global burden, affecting billions and substantially reducing quality of life. While modern therapies ca

DiningBench: A Hierarchical Multi-view Benchmark for Perception and Reasoning in the Dietary Domain

Model ReleasesDGX agent

arXiv:2604.10425v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have revolutionized general visual understanding. However, their application in the food domain rem

Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation

Local AiDGX agent

arXiv:2503.16771v3 Announce Type: replace-cross Abstract: As Large Language Models for Code (LM4Code) become integral to software engineering, establishing trust in their output becomes critical. Howe

Enhancing Fine-Grained Spatial Grounding in 3D CT Report Generation via Discriminative Guidance

SafetyDGX agent

arXiv:2604.10437v1 Announce Type: new Abstract: Vision--language models (VLMs) for radiology report generation (RRG) can produce long-form chest CT reports from volumetric scans and show strong potent

Ernie Image Turbo is Capable of ...

Local AiDGX agent

ERNIE Image Turbo is an open text-to-image generation model developed by Baidu's ERNIE-Image team, serving as the distilled release of the full ERNIE-Image model and built on a single-stream Diffusion

Evaluating Visual Prompts with Eye-Tracking Data for MLLM-Based Human Activity Recognition

ResearchDGX agent

arXiv:2604.09585v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as foundation models for IoT applications such as human activity recognition (HAR). However, directly applyi

GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs

Model ReleasesDGX agent

arXiv:2511.11653v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for passage reranking in information retrieval, leveraging their superior reasonin

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of file…

Model ReleasesDGX agent

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of files that the agent is allowed to edit to hill climb a metric/e

HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement

TutorialsDGX agent

arXiv:2604.10675v1 Announce Type: new Abstract: We propose a method to learn explicit, class-conditioned spatial priors for object placement in natural scenes by distilling the implicit placement know

HiPRAG: Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation

Model ReleasesDGX agent

arXiv:2510.07794v2 Announce Type: replace-cross Abstract: Agentic RAG is a powerful technique for incorporating external information that LLMs lack, enabling better problem solving and question answer

Human vs. Machine Deception: Distinguishing AI-Generated and Human-Written Fake News Using Ensemble Learning

ResearchDGX agent

arXiv:2604.09960v1 Announce Type: new Abstract: The rapid adoption of large language models has introduced a new class of AI-generated fake news that coexists with traditional human-written misinforma

Intent-aligned Formal Specification Synthesis via Traceable Refinement

Model ReleasesDGX agent

arXiv:2604.10392v1 Announce Type: cross Abstract: Large language models are increasingly used to generate code from natural language, but ensuring correctness remains challenging. Formal verification

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs

SafetyDGX agent

arXiv:2604.10403v1 Announce Type: new Abstract: We address jailbreaks, backdoors, and unlearning for large language models (LLMs). Unlike prior work, which trains LLMs based on their actions when give

Learning and Enforcing Context-Sensitive Control for LLMs

TutorialsDGX agent

arXiv:2604.10667v1 Announce Type: cross Abstract: Controlling the output of Large Language Models (LLMs) through context-sensitive constraints has emerged as a promising approach to overcome the limit

Learning to Adapt: In-Context Learning Beyond Stationarity

Model ReleasesDGX agent

arXiv:2604.10946v1 Announce Type: new Abstract: Transformer models have become foundational across a wide range of scientific and engineering domains due to their strong empirical performance. A key c

LLMs Should Incorporate Explicit Mechanisms for Human Empathy

Model ReleasesDGX agent

arXiv:2604.10557v1 Announce Type: cross Abstract: This paper argues that Large Language Models (LLMs) should incorporate explicit mechanisms for human empathy. As LLMs become increasingly deployed in

LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval

Model ReleasesDGX agent

arXiv:2601.14706v3 Announce Type: replace Abstract: In this paper, we present LookBench (We use the term 'look' to reflect retrieval that mirrors how people shop -- finding the exact item, a close sub

LRD-Net: A Lightweight Real-Centered Detection Network for Cross-Domain Face Forgery Detection

Model ReleasesDGX agent

arXiv:2604.10862v1 Announce Type: new Abstract: The rapid advancement of diffusion-based generative models has made face forgery detection a critical challenge in digital forensics. Current detection

MARS: Unleashing the Power of Speculative Decoding via Margin-Aware Verification

Local AiDGX agent

arXiv:2601.15498v2 Announce Type: replace Abstract: Speculative Decoding (SD) accelerates autoregressive large language model (LLM) inference by decoupling generation and verification. While recent me

MERMAID: Memory-Enhanced Retrieval and Reasoning with Multi-Agent Iterative Knowledge Grounding for Veracity Assessment

Model ReleasesDGX agent

arXiv:2601.22361v2 Announce Type: replace-cross Abstract: Assessing the veracity of online content has become increasingly critical. Large language models (LLMs) have recently enabled substantial prog

Min-k Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics

Model ReleasesDGX agent

arXiv:2604.11012v1 Announce Type: new Abstract: The quality of text generated by large language models depends critically on the decoding sampling strategy. While mainstream methods such as Top-k, Top

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark

Model ReleasesDGX agent

arXiv:2604.10755v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced clinical tasks for common conditions, but their performance on rare diseases remains largely unte

MoBo, CPU, RAM suggestion | I'm terrible at guessing hardware | no gpu

HardwareDGX agent

This Reddit thread from r/ollama features a user seeking community recommendations for motherboard, CPU, and RAM components to build a system optimized for running Ollama locally without a dedicated G

Multi-modal, multi-scale representation learning for satellite imagery analysis just needs a good ALiBi

Model ReleasesDGX agent

arXiv:2604.10347v1 Announce Type: new Abstract: Vision foundation models have been shown to be effective at processing satellite imagery into representations fit for downstream tasks, however, creatin

Multimodal Diffusion Forcing for Forceful Manipulation

TutorialsDGX agent

arXiv:2511.04812v2 Announce Type: replace-cross Abstract: Given a dataset of expert trajectories, standard imitation learning approaches typically learn a direct mapping from observations (e.g., RGB i

Near OOD Detection for Vision-Language Prompt Learning with Contrastive Logit Score

ResearchDGX agent

arXiv:2405.16091v2 Announce Type: replace Abstract: Prompt learning has emerged as an efficient and effective method for fine-tuning vision-language models such as CLIP. While many studies have explor

Non-stationary Diffusion For Probabilistic Time Series Forecasting

ApplicationsDGX agent

arXiv:2505.04278v3 Announce Type: replace-cross Abstract: Due to the dynamics of underlying physics and external influences, the uncertainty of time series often varies over time. However, existing De

On The Application of Linear Attention in Multimodal Transformers

ResearchDGX agent

arXiv:2604.10064v1 Announce Type: new Abstract: Multimodal Transformers serve as the backbone for state-of-the-art vision-language models, yet their quadratic attention complexity remains a critical b

Pair2Scene: Learning Local Object Relations for Procedural Scene Generation

Local AiDGX agent

arXiv:2604.11808v1 Announce Type: new Abstract: Generating high-fidelity 3D indoor scenes remains a significant challenge due to data scarcity and the complexity of modeling intricate spatial relation

PEMANT: Persona-Enriched Multi-Agent Negotiation for Travel

SafetyDGX agent

arXiv:2604.10475v1 Announce Type: new Abstract: Modeling household-level trip generation is fundamental to accurate demand forecasting, traffic flow estimation, and urban system planning. Existing stu

Persona Non Grata: Single-Method Safety Evaluation Is Incomplete for Persona-Imbued LLMs

Model ReleasesDGX agent

arXiv:2604.11120v1 Announce Type: new Abstract: Personality imbuing customizes LLM behavior, but safety evaluations almost always study prompt-based personas alone. We show this is incomplete: prompti

Query Lower Bounds for Diffusion Sampling

ResearchDGX agent

arXiv:2604.10857v1 Announce Type: cross Abstract: Diffusion models generate samples by iteratively querying learned score estimates. A rapidly growing literature focuses on accelerating sampling by mi

Resisting Humanization: Ethical Front-End Design Choices in AI for Sensitive Contexts

ApplicationsDGX agent

arXiv:2603.24853v2 Announce Type: replace Abstract: Ethical debates in AI have primarily focused on back-end issues such as data governance, model training, and algorithmic decision-making. Less atten

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

Model ReleasesDGX agent

arXiv:2604.09860v1 Announce Type: cross Abstract: The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid

S4M: 4-points to Segment Anything

ResearchDGX agent

arXiv:2503.05534v3 Announce Type: replace Abstract: Purpose: The Segment Anything Model (SAM) promises to ease the annotation bottleneck in medical segmentation, but overlapping anatomy and blurred bo

SafeConstellations: Mitigating Over-Refusals in LLMs Through Task-Aware Representation Steering

SafetyDGX agent

arXiv:2508.11290v3 Announce Type: replace Abstract: LLMs increasingly exhibit over-refusal behavior, where safety mechanisms cause models to reject benign instructions that seemingly resemble harmful

SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation

HardwareDGX agent

arXiv:2604.10306v1 Announce Type: new Abstract: As Earth-observation workloads move toward onboard and edge processing, remote-sensing segmentation models must operate under tight latency and energy c

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?

Model ReleasesDGX agent

arXiv:2604.10718v1 Announce Type: new Abstract: Accelerating scientific discovery requires the identification of which experiments would yield the best outcomes before committing resources to costly p

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding

Model ReleasesDGX agent

arXiv:2604.11244v1 Announce Type: new Abstract: Advances in Multimodal Large Language Models (MLLMs) are transforming video captioning from a descriptive endpoint into a semantic interface for both vi

SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding

Model ReleasesDGX agent

arXiv:2604.10152v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs)

SPEED-Bench: A Unified and Diverse Benchmark for Speculative Decoding

Model ReleasesDGX agent

arXiv:2604.09557v1 Announce Type: cross Abstract: Speculative Decoding (SD) has emerged as a critical technique for accelerating Large Language Model (LLM) inference. Unlike deterministic system optim

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs

ResearchDGX agent

arXiv:2509.22220v2 Announce Type: replace Abstract: Prevalent semantic speech tokenizers, designed to capture linguistic content, are surprisingly fragile. We find they are not robust to meaning-irrel

STaR-DRO: Stateful Tsallis Reweighting for Group-Robust Structured Prediction

Model ReleasesDGX agent

arXiv:2604.09737v1 Announce Type: cross Abstract: Structured prediction requires models to generate ontology-constrained labels, grounded evidence, and valid structure under ambiguity, label skew, and

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning

TutorialsDGX agent

arXiv:2604.10228v1 Announce Type: new Abstract: Current multimodal models often suffer from shallow reasoning, leading to errors caused by incomplete or inconsistent thought processes. To address this

← Previous
1…369370371372373…1050
Next →