AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Research

DeepWeightFlow: Re-Basined Flow Matching for Generating Neural Network Weights

DGX agent

arXiv:2601.05052v2 Announce Type: replace Abstract: Building efficient and effective generative models for neural network weights has been a research focus of significant interest that faces challenge

researcharxiv-cs-lg
1 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Design Structure Matrix Modularization with Large Language Models

DGX agent

arXiv:2604.28018v1 Announce Type: cross Abstract: Design Structure Matrix (DSM) modularization, the task of partitioning system elements into cohesive modules, is a fundamental combinatorial challenge

safetyarxiv-cs-ai
1 May 2026
Model Releases

Do What I Say: A Spoken Prompt Dataset for Instruction-Following

DGX agent

arXiv:2603.09881v2 Announce Type: replace Abstract: Speech Large Language Models (SLLMs) have rapidly expanded, supporting a wide range of tasks. These models are typically evaluated using text prompt

model-releasesarxiv-cs-cl
1 May 2026
Local Ai

Enhancing Linux Privilege Escalation Attack Capabilities of Local LLM Agents

DGX agent

arXiv:2604.27143v1 Announce Type: cross Abstract: Recent research has demonstrated the potential of Large Language Models (LLMs) for autonomous penetration testing, particularly when using cloud-based

local-aiarxiv-cs-ai
1 May 2026
Research

Fitting Horn DL Ontologies to ABox and Query Examples: A Tale of Simulation Quantifiers and Finite Models

DGX agent

arXiv:2604.26976v1 Announce Type: cross Abstract: We study the problem of fitting a description logic (DL) ontology to a given set of positive and negative examples that take the form of an ABox and a

researcharxiv-cs-ai
1 May 2026
Model Releases

MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction

DGX agent

arXiv:2604.27393v1 Announce Type: new Abstract: Recent progress in multimodal large language models (MLLMs) has brought AI capabilities from static offline data processing to real-time streaming inter

model-releasesarxiv-cs-cl
1 May 2026
Research

Modeling Spatial Extremal Dependence of Precipitation Using Distributional Neural Networks

DGX agent

arXiv:2407.08668v3 Announce Type: replace-cross Abstract: In this work, we propose a simulation-based estimation approach using generative neural networks to determine dependencies of precipitation ma

researcharxiv-cs-lg
1 May 2026
Research

Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading

DGX agent

arXiv:2604.27637v1 Announce Type: new Abstract: Current Large Language Model (LLM) evaluation frameworks utilize the same static prompt template across all models under evaluation. This differs from t

researcharxiv-cs-ai
1 May 2026
Model Releases

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs

DGX agent

arXiv:2604.27401v1 Announce Type: new Abstract: Perturbation probing generates task-specific causal hypotheses for FFN neurons in large language models using two forward passes per prompt and no backp

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

PVeRA: Probabilistic Vector-Based Random Matrix Adaptation

DGX agent

arXiv:2512.07703v2 Announce Type: replace Abstract: Large foundation models have emerged in the last years and are pushing performance boundaries for a variety of tasks. Training or even finetuning su

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension

DGX agent

arXiv:2601.14289v2 Announce Type: replace-cross Abstract: Understanding research papers remains challenging for foundation models due to specialized scientific discourse and complex figures and tables

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

RuC: HDL-Agnostic Rule Completion Benchmark Generation

DGX agent

arXiv:2604.27780v1 Announce Type: cross Abstract: Large Language Models (LLMs) have rapidly improved in performance across code-related tasks, making their integration into Register Transfer Level (RT

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images

DGX agent

arXiv:2604.28039v1 Announce Type: new Abstract: Spectra are a prevalent yet highly information-dense form of scientific imagery, presenting substantial challenges to multimodal large language models (

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

A Practice of Post-Training on Llama-3 70B with Optimal Selection of Additional Language Mixture Ratio

DGX agent

arXiv:2409.06624v4 Announce Type: replace-cross Abstract: Large Language Models (LLM) often need to be Continual Pre-Trained (CPT) to obtain unfamiliar language skills or adapt to new domains. The hug

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

A Systematic Comparison of Prompting and Multi-Agent Methods for LLM-based Stance Detection

DGX agent

arXiv:2604.26319v1 Announce Type: new Abstract: Stance detection identifies the attitude of a text author toward a given target. Recent studies have explored various LLM-based strategies for this task

model-releasesarxiv-cs-cl
30 Apr 2026
Research

Budget-Constrained Causal Bandits: Bridging Uplift Modeling and Sequential Decision-Making

DGX agent

arXiv:2604.26169v1 Announce Type: new Abstract: Treatment allocation under budget constraints is a central challenge in digital advertising: advertisers must decide which users to show ads to while sp

researcharxiv-cs-lg
30 Apr 2026
Model Releases

CoQuant: Joint Weight-Activation Subspace Projection for Mixed-Precision LLMs

DGX agent

arXiv:2604.26378v1 Announce Type: new Abstract: Post-training quantization (PTQ) has become an important technique for reducing the inference cost of Large Language Models (LLMs). While recent mixed-p

model-releasesarxiv-cs-lg
30 Apr 2026
Local Ai

Emergent Coordination in Multi-Agent Language Models

DGX agent

arXiv:2510.05174v4 Announce Type: replace-cross Abstract: When are multi-agent LLM systems merely a collection of individual agents versus an integrated collective with higher-order structure? We intr

local-aiarxiv-cs-ai
30 Apr 2026
Model Releases

MoRFI: Monotonic Sparse Autoencoder Feature Identification

DGX agent

arXiv:2604.26866v1 Announce Type: new Abstract: Large language models (LLMs) acquire most of their factual knowledge during the pre-training stage, through next token prediction. Subsequent stages of

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Thinking with Drafting: Optical Decompression via Logical Reconstruction

DGX agent

arXiv:2602.11731v2 Announce Type: replace Abstract: Existing multimodal large language models have achieved high-fidelity visual perception and exploratory visual generation. However, a precision para

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

VulStyle: A Multi-Modal Pre-Training for Code Stylometry-Augmented Vulnerability Detection

DGX agent

arXiv:2604.26313v1 Announce Type: cross Abstract: We present VulStyle, a multi-modal software vulnerability detection model that jointly encodes function-level source code, non-terminal Abstract Synta

model-releasesarxiv-cs-lg
30 Apr 2026
Safety

When Annotators Disagree, Topology Explains: Mapper, a Topological Tool for Exploring Text Embedding Geometry and Ambiguity

DGX agent

arXiv:2510.17548v2 Announce Type: replace Abstract: Language models are often evaluated with scalar metrics like accuracy, but such measures fail to capture how models internally represent ambiguity,

safetyarxiv-cs-cl
30 Apr 2026
Model Releases

Below-Chance Blindness: Prompted Underperformance in Small LLMs Produces Positional Bias Rather than Answer Avoidance

DGX agent

arXiv:2604.25249v1 Announce Type: new Abstract: Detecting sandbagging--the deliberate underperformance on capability evaluations--is an open problem in AI safety. We tested whether symptom validity te

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Evaluating LLM Safety Under Repeated Inference via Accelerated Prompt Stress Testing

DGX agent

arXiv:2602.11786v2 Announce Type: replace Abstract: Traditional benchmarks for large language models (LLMs), such as HELM and AIR-BENCH, primarily assess safety through breadth-oriented evaluation acr

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

FARM: Enhancing Molecular Representations with Functional Group Awareness

DGX agent

arXiv:2410.02082v4 Announce Type: replace Abstract: We introduce Functional Group-Aware Representations for Small Molecules (FARM), a novel foundation model designed to bridge the gap between SMILES,

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

The Thinking Pixel: Recursive Sparse Reasoning in Multimodal Diffusion Latents

DGX agent

arXiv:2604.25299v1 Announce Type: new Abstract: Diffusion models have achieved success in high-fidelity data synthesis, yet their capacity for more complex, structured reasoning like text following ta

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Toward Multimodal Conversational AI for Age-Related Macular Degeneration

DGX agent

arXiv:2604.25720v1 Announce Type: cross Abstract: Despite strong performance of deep learning models in retinal disease detection, most systems produce static predictions without clinical reasoning or

researcharxiv-cs-cl
29 Apr 2026
Safety

Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation

DGX agent

arXiv:2511.21517v2 Announce Type: replace Abstract: Unlike text, speech conveys information about the speaker, such as gender, through acoustic cues like pitch. This gives rise to modality-specific bi

safetyarxiv-cs-cl
29 Apr 2026
Safety

A Comparative analysis of Layer-wise Representational Capacity in AR and Diffusion LLMs

DGX agent

arXiv:2603.07475v2 Announce Type: replace Abstract: Autoregressive (AR) language models build representations incrementally via left-to-right prediction, while diffusion language models (dLLMs) are tr

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

AeSlides: Incentivizing Aesthetic Layout in LLM-Based Slide Generation via Verifiable Rewards

DGX agent

arXiv:2604.22840v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong potential in agentic tasks, particularly in slide generation. However, slide generation poses a

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Agri-CPJ: A Training-Free Explainable Framework for Agricultural Pest Diagnosis Using Caption-Prompt-Judge and LLM-as-a-Judge

DGX agent

arXiv:2604.23701v1 Announce Type: cross Abstract: Crop disease diagnosis from field photographs faces two recurring problems: models that score well on benchmarks frequently hallucinate species names,

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

AI Safety Training Can be Clinically Harmful

DGX agent

arXiv:2604.23445v1 Announce Type: cross Abstract: Large language models are being deployed as mental health support agents at scale, yet only 16% of LLM-based chatbot interventions have undergone rigo

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code

DGX agent

arXiv:2604.23361v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong performance on a wide range of software engineering tasks, including code generation and analysi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

C-MORAL: Controllable Multi-Objective Molecular Optimization with Reinforcement Alignment for LLMs

DGX agent

arXiv:2604.23061v1 Announce Type: cross Abstract: Large language models (LLMs) show promise for molecular optimization, but aligning them with selective and competing drug-design constraints remains c

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Clotho: Measuring Task-Specific Pre-Generation Test Adequacy for LLM Inputs

DGX agent

arXiv:2509.17314v3 Announce Type: replace-cross Abstract: Software increasingly relies on the emergent capabilities of Large Language Models (LLMs), from natural language understanding to program anal

model-releasesarxiv-cs-lg
28 Apr 2026
Safety

Context-Aware Hospitalization Forecasting Evaluations for Decision Support using LLMs

DGX agent

arXiv:2604.23949v1 Announce Type: new Abstract: Medical and public health experts must make real-time resource decisions, such as expanding hospital bed capacity, based on projected hospitalization tr

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Don't Make the LLM Read the Graph: Make the Graph Think

DGX agent

arXiv:2604.23057v1 Announce Type: new Abstract: We investigate whether explicit belief graphs improve LLM performance in cooperative multi-agent reasoning. Through 3,000+ controlled trials across four

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

DRIFT: Transferring Reasoning Priors for Efficient MLLM Fine-Tuning

DGX agent

arXiv:2510.15050v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made rapid progress, yet their reasoning ability often lags behind strong text-only LLMs. Bridging thi

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems

DGX agent

arXiv:2604.24477v1 Announce Type: cross Abstract: The rapid integration of Large Language Models (LLMs) into Multi-Agent Systems (MAS) has significantly enhanced their collaborative problem-solving ca

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Guided Speculative Inference for Efficient Test-Time Alignment of LLMs

DGX agent

arXiv:2506.04118v3 Announce Type: replace Abstract: We propose Guided Speculative Inference (GSI), a novel algorithm for efficient reward-guided decoding in large language models. GSI combines soft be

safetyarxiv-cs-lg
28 Apr 2026
Research

Knowledge Vector of Logical Reasoning in Large Language Models

DGX agent

arXiv:2604.23877v1 Announce Type: new Abstract: Logical reasoning serve as a central capability in LLMs and includes three main forms: deductive, inductive, and abductive reasoning. In this work, we s

researcharxiv-cs-cl
28 Apr 2026
Research

LILogic Net: Compact Logic Gate Networks with Learnable Connectivity for Efficient Hardware Deployment

DGX agent

arXiv:2511.12340v2 Announce Type: replace Abstract: Efficient machine learning deployment requires models that account for hardware constraints. Because binary logic gates are the fundamental primitiv

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Long-Context Aware Upcycling: A New Frontier for Hybrid LLM Scaling

DGX agent

arXiv:2604.24715v1 Announce Type: new Abstract: Hybrid sequence models that combine efficient Transformer components with linear sequence modeling blocks are a promising alternative to pure Transforme

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation

DGX agent

arXiv:2511.14967v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown great promise in generating structured diagrams from natural language descriptions, particularly Merma

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Meta-CoT: Enhancing Granularity and Generalization in Image Editing

DGX agent

arXiv:2604.24625v1 Announce Type: cross Abstract: Unified multi-modal understanding/generative models have shown improved image editing performance by incorporating fine-grained understanding into the

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ServImage: An Image Generation and Editing Benchmark from Real-world Commercial Imaging Services

DGX agent

arXiv:2604.24023v1 Announce Type: new Abstract: Recent image generation and editing models demonstrate robust adherence to instructions and high visual quality on academic benchmarks. However, their p

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Sphere-Depth: A Benchmark for Depth Estimation Methods with Varying Spherical Camera Orientations

DGX agent

arXiv:2604.23432v1 Announce Type: cross Abstract: Reliable depth estimation from spherical images is crucial for 360{eg} vision in robotic navigation and immersive scene understanding. However, the on

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

The Pragmatic Persona: Discovering LLM Persona through Bridging Inference

DGX agent

arXiv:2604.24079v1 Announce Type: cross Abstract: Large Language Models (LLMs) reveal inherent and distinctive personas through dialogue. However, most existing persona discovery approaches rely on su

model-releasesarxiv-cs-ai
28 Apr 2026
← Previous
1…293294295296297…1058
Next →