AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Safety

OPSDL: On-Policy Self-Distillation for Long-Context Language Models

DGX agent

arXiv:2604.17535v1 Announce Type: new Abstract: Extending the effective context length of large language models (LLMs) remains a central challenge for real-world applications. While recent post-traini

safetyarxiv-cs-cl
21 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Parallel Test-Time Scaling for Latent Reasoning Models

DGX agent

arXiv:2510.07745v4 Announce Type: replace Abstract: Parallel test-time scaling (TTS) is a pivotal approach for enhancing large language models (LLMs), typically by sampling multiple token-based chains

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not

DGX agent

arXiv:2604.04825v2 Announce Type: replace Abstract: Large language models achieve strong performance on many language tasks, yet it remains unclear whether they integrate world knowledge with syntacti

tutorialsarxiv-cs-cl
21 Apr 2026
Safety

Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models

DGX agent

arXiv:2601.15220v2 Announce Type: replace Abstract: We identify a novel phenomenon in language models: benign fine-tuning of frontier models can lead to privacy collapse. We find that diverse, subtle

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Program Structure-aware Language Models: Targeted Software Testing beyond Textual Semantics

DGX agent

arXiv:2604.17715v1 Announce Type: cross Abstract: Recent advances in large language models for test case generation have improved branch coverage via prompt-engineered mutations. However, they still l

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Residual Diffusion Bridge Model for Image Restoration

DGX agent

arXiv:2510.23116v3 Announce Type: replace Abstract: Diffusion bridge models establish probabilistic paths between arbitrary paired distributions and exhibit great potential for universal image restora

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Revisiting Change VQA in Remote Sensing with Structured and Native Multimodal Qwen Models

DGX agent

arXiv:2604.18429v1 Announce Type: new Abstract: Change visual question answering (Change VQA) addresses the problem of answering natural-language questions about semantic changes between bi-temporal r

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

S2H-DPO: Hardness-Aware Preference Optimization for Vision-Language Models

DGX agent

arXiv:2604.18512v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in single-image understanding, yet effective reasoning across multiple images remain

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Sky2Ground: A Benchmark for Site Modeling under Varying Altitude

DGX agent

arXiv:2603.13740v3 Announce Type: replace Abstract: We introduce Sky2Ground, a three-view dataset designed for varying altitude camera localization, correspondence learning, and reconstruction. The da

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Test-Time Adaptation for EEG Foundation Models: A Systematic Study under Real-World Distribution Shifts

DGX agent

arXiv:2604.16926v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models have shown strong potential for learning generalizable representations from large-scale neural data, yet

model-releasesarxiv-cs-lg
21 Apr 2026
Research

When Background Matters: Breaking Medical Vision Language Models by Transferable Attack

DGX agent

arXiv:2604.17318v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly used in clinical diagnostics, yet their robustness to adversarial attacks remains largely unexplored, pos

researcharxiv-cs-cv
21 Apr 2026
Safety

When Choices Become Risks: Safety Failures of Large Language Models under Multiple-Choice Constraints

DGX agent

arXiv:2604.16916v1 Announce Type: new Abstract: Safety alignment in large language models (LLMs) is primarily evaluated under open-ended generation, where models can mitigate risk by refusing to respo

safetyarxiv-cs-cl
21 Apr 2026
Research

When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models

DGX agent

arXiv:2507.13868v2 Announce Type: replace Abstract: Vision-language models (VLMs) increasingly combine visual and textual information to perform complex tasks. However, conflicts between their interna

researcharxiv-cs-cv
21 Apr 2026
Local Ai

AdaVFM: Adaptive Vision Foundation Models for Edge Intelligence via LLM-Guided Execution

DGX agent

arXiv:2604.15622v1 Announce Type: new Abstract: Language-aligned vision foundation models (VFMs) enable versatile visual understanding for always-on contextual AI, but their deployment on edge devices

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

ChatENV: An Interactive Vision-Language Model for Sensor-Guided Environmental Monitoring and Scenario Simulation

DGX agent

arXiv:2508.10635v3 Announce Type: replace Abstract: Understanding environmental changes from remote sensing imagery is vital for climate resilience, urban planning, and ecosystem monitoring. Yet, curr

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

FineSteer: A Unified Framework for Fine-Grained Inference-Time Steering in Large Language Models

DGX agent

arXiv:2604.15488v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit undesirable behaviors, such as safety violations and hallucinations. Although inference-time steering offer

safetyarxiv-cs-ai
20 Apr 2026
Research

How Hypocritical Is Your LLM judge? Listener-Speaker Asymmetries in the Pragmatic Competence of Large Language Models

DGX agent

arXiv:2604.15873v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly studied as repositories of linguistic knowledge. In this line of work, models are commonly evaluated both

researcharxiv-cs-cl
20 Apr 2026
Safety

Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information

DGX agent

arXiv:2604.15701v1 Announce Type: new Abstract: The significant computational demands of large language models have increased interest in distilling reasoning abilities into smaller models via Chain-o

safetyarxiv-cs-cl
20 Apr 2026
Model Releases

Preference Estimation via Opponent Modeling in Multi-Agent Negotiation

DGX agent

arXiv:2604.15687v1 Announce Type: new Abstract: Automated negotiation in complex, multi-party and multi-issue settings critically depends on accurate opponent modeling. However, conventional numerical

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Repurposing 3D Generative Model for Autoregressive Layout Generation

DGX agent

arXiv:2604.16299v1 Announce Type: new Abstract: We introduce LaviGen, a framework that repurposes 3D generative models for 3D layout generation. Unlike previous methods that infer object layouts from

model-releasesarxiv-cs-cv
20 Apr 2026
Applications

Tabular foundation models for in-context prediction of molecular properties

DGX agent

arXiv:2604.16123v1 Announce Type: new Abstract: Accurate molecular property prediction is central to drug discovery, catalysis, and process design, yet real-world applications are often limited by sma

applicationsarxiv-cs-lg
20 Apr 2026
Research

Assessing the Performance-Efficiency Trade-off of Foundation Models in Probabilistic Electricity Price Forecasting

DGX agent

arXiv:2604.14739v1 Announce Type: new Abstract: Large-scale renewable energy deployment introduces pronounced volatility into the electricity system, turning grid operation into a complex stochastic o

researcharxiv-cs-lg
17 Apr 2026
Model Releases

CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification

DGX agent

arXiv:2604.14602v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate toxic content, posing significant risks for safe deployment. Current mitigation strategies often degrad

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

CobwebTM: Probabilistic Concept Formation for Lifelong and Hierarchical Topic Modeling

DGX agent

arXiv:2604.14489v1 Announce Type: new Abstract: Topic modeling seeks to uncover latent semantic structure in text corpora with minimal supervision. Neural approaches achieve strong performance but req

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

DharmaOCR: Specialized Small Language Models for Structured OCR that outperform Open-Source and Commercial Baselines

DGX agent

arXiv:2604.14314v1 Announce Type: cross Abstract: This manuscript introduces DharmaOCR Full and Lite, a pair of specialized small language models (SSLMs) for structured OCR that jointly optimize trans

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Domain-Adaptive Model Merging Across Disconnected Modes

DGX agent

arXiv:2603.05957v2 Announce Type: replace-cross Abstract: Learning across domains is challenging when data cannot be centralized due to privacy or heterogeneity, which limits the ability to train a si

researcharxiv-cs-ai
17 Apr 2026
Research

Exploring the flavor structure of leptons via diffusion models

DGX agent

arXiv:2503.21432v2 Announce Type: replace-cross Abstract: We propose a method to explore the flavor structure of leptons using diffusion models, which are known as one of generative artificial intelli

researcharxiv-cs-lg
17 Apr 2026
Safety

Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI

DGX agent

arXiv:2403.10559v3 Announce Type: replace Abstract: This report investigates the history and impact of Generative Models and Connected and Automated Vehicles (CAVs), two groundbreaking forces pushing

safetyarxiv-cs-lg
17 Apr 2026
Research

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds

DGX agent

arXiv:2604.14268v1 Announce Type: new Abstract: We introduce HY-World 2.0, a multi-modal world model framework that advances our prior project HY-World 1.0. HY-World 2.0 accommodates diverse input mod

researcharxiv-cs-cv
17 Apr 2026
Model Releases

Prompt-Guided Image Editing with Masked Logit Nudging in Visual Autoregressive Models

DGX agent

arXiv:2604.14591v1 Announce Type: new Abstract: We address the problem of prompt-guided image editing in visual autoregressive models. Given a source image and a target text prompt, we aim to modify t

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

ProRank: Prompt Warmup via Reinforcement Learning for Small Language Models Reranking

DGX agent

arXiv:2506.03487v3 Announce Type: replace-cross Abstract: Reranking is fundamental to information retrieval and retrieval-augmented generation, with recent Large Language Models (LLMs) significantly a

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Psychological Steering of Large Language Models

DGX agent

arXiv:2604.14463v1 Announce Type: new Abstract: Large language models (LLMs) emulate a consistent human-like behavior that can be shaped through activation-level interventions. This paradigm is conver

researcharxiv-cs-cl
17 Apr 2026
Model Releases

QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading Strategies

DGX agent

arXiv:2604.15151v1 Announce Type: new Abstract: Large language models have demonstrated strong performance on general-purpose programming tasks, yet their ability to generate executable algorithmic tr

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Style Amnesia: Investigating Speaking Style Degradation and Mitigation in Multi-Turn Spoken Language Models

DGX agent

arXiv:2512.23578v3 Announce Type: replace Abstract: In this paper, we show that when spoken language models (SLMs) are instructed to speak in a specific speaking style at the beginning of a multi-turn

researcharxiv-cs-cl
17 Apr 2026
Model Releases

fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI Understanding

DGX agent

arXiv:2511.21760v3 Announce Type: replace Abstract: Recent advances in multimodal large language models (LLMs) have enabled unified reasoning across images, audio, and video, but extending such capabi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

L2D-Clinical: Learning to Defer for Adaptive Model Selection in Clinical Text Classification

DGX agent

arXiv:2604.13285v1 Announce Type: new Abstract: Clinical text classification requires choosing between specialized fine-tuned models (BERT variants) and general-purpose large language models (LLMs), y

model-releasesarxiv-cs-cl
16 Apr 2026
Research

ReproMIA: A Comprehensive Analysis of Model Reprogramming for Proactive Membership Inference Attacks

DGX agent

arXiv:2603.28942v3 Announce Type: replace Abstract: The pervasive deployment of deep learning models across critical domains has concurrently intensified privacy concerns due to their inherent propens

researcharxiv-cs-lg
16 Apr 2026
Model Releases

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges

DGX agent

arXiv:2604.13602v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) and related alignment paradigms have become central to steering large language models (LLMs) and multi

model-releasesarxiv-cs-lg
16 Apr 2026
Local Ai

Simplicity Prevails: The Emergence of Generalizable AIGI Detection in Visual Foundation Models

DGX agent

arXiv:2602.01738v2 Announce Type: replace Abstract: While specialized detectors for AI-Generated Images (AIGI) achieve near-perfect accuracy on curated benchmarks, they suffer from a dramatic performa

local-aiarxiv-cs-cv
16 Apr 2026
Safety

UNBOX: Unveiling Black-box visual models with Natural-language

DGX agent

arXiv:2603.08639v2 Announce Type: replace Abstract: Ensuring trustworthiness in open-world visual recognition requires models that are interpretable, fair, and robust to distribution shifts. Yet moder

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

A Foot Resistive Force Model for Legged Locomotion on Muddy Terrains

DGX agent

arXiv:2604.12006v1 Announce Type: new Abstract: Legged robots face significant challenges in moving and navigating on deformable and highly yielding terrain such as mud. We present a resistive force m

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks

DGX agent

arXiv:2604.12379v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly rely on explicit reasoning to solve coding tasks, yet evaluating the quality of this reasoning remains chall

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Bipedal-Walking-Dynamics Model on Granular Terrains

DGX agent

arXiv:2604.11981v1 Announce Type: new Abstract: Bipeds have demonstrated high agility and mobility in unstructured environments such as sand. The yielding of such granular media brings significant sin

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

Fragile Preferences: A Deep Dive Into Order Effects in Large Language Models

DGX agent

arXiv:2506.14092v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in decision-support systems for high-stakes domains such as hiring and university admissions,

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Fully Homomorphic Encryption on Llama 3 model for privacy preserving LLM inference

DGX agent

arXiv:2604.12168v1 Announce Type: cross Abstract: The applications of Generative Artificial Intelligence (GenAI) and their intersections with data-driven fields, such as healthcare, finance, transport

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Hard Negative Sample-Augmented DPO Post-Training for Small Language Models

DGX agent

arXiv:2512.19728v2 Announce Type: replace Abstract: Large language models (LLMs) continue to struggle with mathematical reasoning, and common post-training pipelines often reduce each generated soluti

model-releasesarxiv-cs-lg
15 Apr 2026
Local Ai

HintMR: Eliciting Stronger Mathematical Reasoning in Small Language Models

DGX agent

arXiv:2604.12229v1 Announce Type: new Abstract: Small language models (SLMs) often struggle with complex mathematical reasoning due to limited capacity to maintain long chains of intermediate steps an

local-aiarxiv-cs-ai
15 Apr 2026
Model Releases

Mining Large Language Models for Low-Resource Language Data: Comparing Elicitation Strategies for Hausa and Fongbe

DGX agent

arXiv:2604.12477v1 Announce Type: cross Abstract: Large language models (LLMs) are trained on data contributed by low-resource language communities, yet the linguistic knowledge encoded in these model

model-releasesarxiv-cs-ai
15 Apr 2026
← Previous
1…5455565758…1021
Next →