AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of file…

DGX agent

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of files that the agent is allowed to edit to hill climb a metric/e

model-releasesharrison-chase--x
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement

DGX agent

arXiv:2604.10675v1 Announce Type: new Abstract: We propose a method to learn explicit, class-conditioned spatial priors for object placement in natural scenes by distilling the implicit placement know

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

HiPRAG: Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation

DGX agent

arXiv:2510.07794v2 Announce Type: replace-cross Abstract: Agentic RAG is a powerful technique for incorporating external information that LLMs lack, enabling better problem solving and question answer

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Human vs. Machine Deception: Distinguishing AI-Generated and Human-Written Fake News Using Ensemble Learning

DGX agent

arXiv:2604.09960v1 Announce Type: new Abstract: The rapid adoption of large language models has introduced a new class of AI-generated fake news that coexists with traditional human-written misinforma

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Intent-aligned Formal Specification Synthesis via Traceable Refinement

DGX agent

arXiv:2604.10392v1 Announce Type: cross Abstract: Large language models are increasingly used to generate code from natural language, but ensuring correctness remains challenging. Formal verification

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs

DGX agent

arXiv:2604.10403v1 Announce Type: new Abstract: We address jailbreaks, backdoors, and unlearning for large language models (LLMs). Unlike prior work, which trains LLMs based on their actions when give

safetyarxiv-cs-lg
14 Apr 2026
Tutorials

Learning and Enforcing Context-Sensitive Control for LLMs

DGX agent

arXiv:2604.10667v1 Announce Type: cross Abstract: Controlling the output of Large Language Models (LLMs) through context-sensitive constraints has emerged as a promising approach to overcome the limit

tutorialsarxiv-cs-ai
14 Apr 2026
Model Releases

Learning to Adapt: In-Context Learning Beyond Stationarity

DGX agent

arXiv:2604.10946v1 Announce Type: new Abstract: Transformer models have become foundational across a wide range of scientific and engineering domains due to their strong empirical performance. A key c

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

LLMs Should Incorporate Explicit Mechanisms for Human Empathy

DGX agent

arXiv:2604.10557v1 Announce Type: cross Abstract: This paper argues that Large Language Models (LLMs) should incorporate explicit mechanisms for human empathy. As LLMs become increasingly deployed in

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval

DGX agent

arXiv:2601.14706v3 Announce Type: replace Abstract: In this paper, we present LookBench (We use the term 'look' to reflect retrieval that mirrors how people shop -- finding the exact item, a close sub

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LRD-Net: A Lightweight Real-Centered Detection Network for Cross-Domain Face Forgery Detection

DGX agent

arXiv:2604.10862v1 Announce Type: new Abstract: The rapid advancement of diffusion-based generative models has made face forgery detection a critical challenge in digital forensics. Current detection

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

MARS: Unleashing the Power of Speculative Decoding via Margin-Aware Verification

DGX agent

arXiv:2601.15498v2 Announce Type: replace Abstract: Speculative Decoding (SD) accelerates autoregressive large language model (LLM) inference by decoupling generation and verification. While recent me

local-aiarxiv-cs-lg
14 Apr 2026
Model Releases

MERMAID: Memory-Enhanced Retrieval and Reasoning with Multi-Agent Iterative Knowledge Grounding for Veracity Assessment

DGX agent

arXiv:2601.22361v2 Announce Type: replace-cross Abstract: Assessing the veracity of online content has become increasingly critical. Large language models (LLMs) have recently enabled substantial prog

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Min-k Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics

DGX agent

arXiv:2604.11012v1 Announce Type: new Abstract: The quality of text generated by large language models depends critically on the decoding sampling strategy. While mainstream methods such as Top-k, Top

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark

DGX agent

arXiv:2604.10755v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced clinical tasks for common conditions, but their performance on rare diseases remains largely unte

model-releasesarxiv-cs-cv
14 Apr 2026
Hardware

MoBo, CPU, RAM suggestion | I'm terrible at guessing hardware | no gpu

DGX agent

This Reddit thread from r/ollama features a user seeking community recommendations for motherboard, CPU, and RAM components to build a system optimized for running Ollama locally without a dedicated G

hardwarer-ollama
14 Apr 2026
Model Releases

Multi-modal, multi-scale representation learning for satellite imagery analysis just needs a good ALiBi

DGX agent

arXiv:2604.10347v1 Announce Type: new Abstract: Vision foundation models have been shown to be effective at processing satellite imagery into representations fit for downstream tasks, however, creatin

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

Multimodal Diffusion Forcing for Forceful Manipulation

DGX agent

arXiv:2511.04812v2 Announce Type: replace-cross Abstract: Given a dataset of expert trajectories, standard imitation learning approaches typically learn a direct mapping from observations (e.g., RGB i

tutorialsarxiv-cs-ai
14 Apr 2026
Research

Near OOD Detection for Vision-Language Prompt Learning with Contrastive Logit Score

DGX agent

arXiv:2405.16091v2 Announce Type: replace Abstract: Prompt learning has emerged as an efficient and effective method for fine-tuning vision-language models such as CLIP. While many studies have explor

researcharxiv-cs-cv
14 Apr 2026
Applications

Non-stationary Diffusion For Probabilistic Time Series Forecasting

DGX agent

arXiv:2505.04278v3 Announce Type: replace-cross Abstract: Due to the dynamics of underlying physics and external influences, the uncertainty of time series often varies over time. However, existing De

applicationsarxiv-cs-ai
14 Apr 2026
Research

On The Application of Linear Attention in Multimodal Transformers

DGX agent

arXiv:2604.10064v1 Announce Type: new Abstract: Multimodal Transformers serve as the backbone for state-of-the-art vision-language models, yet their quadratic attention complexity remains a critical b

researcharxiv-cs-cv
14 Apr 2026
Local Ai

Pair2Scene: Learning Local Object Relations for Procedural Scene Generation

DGX agent

arXiv:2604.11808v1 Announce Type: new Abstract: Generating high-fidelity 3D indoor scenes remains a significant challenge due to data scarcity and the complexity of modeling intricate spatial relation

local-aiarxiv-cs-cv
14 Apr 2026
Safety

PEMANT: Persona-Enriched Multi-Agent Negotiation for Travel

DGX agent

arXiv:2604.10475v1 Announce Type: new Abstract: Modeling household-level trip generation is fundamental to accurate demand forecasting, traffic flow estimation, and urban system planning. Existing stu

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Persona Non Grata: Single-Method Safety Evaluation Is Incomplete for Persona-Imbued LLMs

DGX agent

arXiv:2604.11120v1 Announce Type: new Abstract: Personality imbuing customizes LLM behavior, but safety evaluations almost always study prompt-based personas alone. We show this is incomplete: prompti

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Query Lower Bounds for Diffusion Sampling

DGX agent

arXiv:2604.10857v1 Announce Type: cross Abstract: Diffusion models generate samples by iteratively querying learned score estimates. A rapidly growing literature focuses on accelerating sampling by mi

researcharxiv-cs-ai
14 Apr 2026
Applications

Resisting Humanization: Ethical Front-End Design Choices in AI for Sensitive Contexts

DGX agent

arXiv:2603.24853v2 Announce Type: replace Abstract: Ethical debates in AI have primarily focused on back-end issues such as data governance, model training, and algorithmic decision-making. Less atten

applicationsarxiv-cs-ai
14 Apr 2026
Model Releases

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

DGX agent

arXiv:2604.09860v1 Announce Type: cross Abstract: The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid

model-releasesarxiv-cs-ai
14 Apr 2026
Research

S4M: 4-points to Segment Anything

DGX agent

arXiv:2503.05534v3 Announce Type: replace Abstract: Purpose: The Segment Anything Model (SAM) promises to ease the annotation bottleneck in medical segmentation, but overlapping anatomy and blurred bo

researcharxiv-cs-cv
14 Apr 2026
Safety

SafeConstellations: Mitigating Over-Refusals in LLMs Through Task-Aware Representation Steering

DGX agent

arXiv:2508.11290v3 Announce Type: replace Abstract: LLMs increasingly exhibit over-refusal behavior, where safety mechanisms cause models to reject benign instructions that seemingly resemble harmful

safetyarxiv-cs-cl
14 Apr 2026
Hardware

SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation

DGX agent

arXiv:2604.10306v1 Announce Type: new Abstract: As Earth-observation workloads move toward onboard and edge processing, remote-sensing segmentation models must operate under tight latency and energy c

hardwarearxiv-cs-cv
14 Apr 2026
Model Releases

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?

DGX agent

arXiv:2604.10718v1 Announce Type: new Abstract: Accelerating scientific discovery requires the identification of which experiments would yield the best outcomes before committing resources to costly p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding

DGX agent

arXiv:2604.11244v1 Announce Type: new Abstract: Advances in Multimodal Large Language Models (MLLMs) are transforming video captioning from a descriptive endpoint into a semantic interface for both vi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding

DGX agent

arXiv:2604.10152v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs)

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SPEED-Bench: A Unified and Diverse Benchmark for Speculative Decoding

DGX agent

arXiv:2604.09557v1 Announce Type: cross Abstract: Speculative Decoding (SD) has emerged as a critical technique for accelerating Large Language Model (LLM) inference. Unlike deterministic system optim

model-releasesarxiv-cs-ai
14 Apr 2026
Research

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs

DGX agent

arXiv:2509.22220v2 Announce Type: replace Abstract: Prevalent semantic speech tokenizers, designed to capture linguistic content, are surprisingly fragile. We find they are not robust to meaning-irrel

researcharxiv-cs-cl
14 Apr 2026
Model Releases

STaR-DRO: Stateful Tsallis Reweighting for Group-Robust Structured Prediction

DGX agent

arXiv:2604.09737v1 Announce Type: cross Abstract: Structured prediction requires models to generate ontology-constrained labels, grounded evidence, and valid structure under ambiguity, label skew, and

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning

DGX agent

arXiv:2604.10228v1 Announce Type: new Abstract: Current multimodal models often suffer from shallow reasoning, leading to errors caused by incomplete or inconsistent thought processes. To address this

tutorialsarxiv-cs-ai
14 Apr 2026
Research

The Myth of Expert Specialization in MoEs: Why Routing Reflects Geometry, Not Necessarily Domain Expertise

DGX agent

arXiv:2604.09780v1 Announce Type: new Abstract: Mixture of Experts (MoEs) are now ubiquitous in large language models, yet the mechanisms behind their 'expert specialization' remain poorly understood.

researcharxiv-cs-ai
14 Apr 2026
Applications

(This is a big part of what was called emergence in earlier academic work on unexpected LLM ability gains)

DGX agent

Ethan Mollick discusses the concept of 'emergence' in large language models (LLMs), referring to the phenomenon where AI systems appear to suddenly develop unexpected capabilities as they scale. The p

applicationsethan-mollick--x
14 Apr 2026
Model Releases

Toward Generalized Cross-Lingual Hateful Language Detection with Web-Scale Data and Ensemble LLM Annotations

DGX agent

arXiv:2604.09625v1 Announce Type: new Abstract: We study whether large-scale unlabelled web data and LLM-based synthetic annotations can improve multilingual hate speech detection. Starting from texts

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Trajectory-based actuator identification via differentiable simulation

DGX agent

arXiv:2604.10351v1 Announce Type: new Abstract: Accurate actuation models are critical for bridging the gap between simulation and real robot behavior, yet obtaining high-fidelity actuator dynamics ty

safetyarxiv-cs-ro
14 Apr 2026
Model Releases

Trust Your Memory: Verifiable Control of Smart Homes through Reinforcement Learning with Multi-dimensional Rewards

DGX agent

arXiv:2604.10110v1 Announce Type: new Abstract: Large Language Models (LLMs) have become a key foundation for enabling personalized smart home experiences. While existing studies have explored how sma

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Tuning Qwen2.5-VL to Improve Its Web Interaction Skills

DGX agent

arXiv:2604.09571v1 Announce Type: cross Abstract: Recent advances in vision-language models (VLMs) have sparked growing interest in using them to automate web tasks, yet their feasibility as independe

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Turning Generators into Retrievers: Unlocking MLLMs for Natural Language-Guided Geo-Localization

DGX agent

arXiv:2604.10721v1 Announce Type: cross Abstract: Natural-language Guided Cross-view Geo-localization (NGCG) aims to retrieve geo-tagged satellite imagery using textual descriptions of ground scenes.

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Variable Selection Using Relative Importance Rankings

DGX agent

arXiv:2509.10853v2 Announce Type: replace-cross Abstract: Although conceptually related, variable selection and relative importance (RI) analysis have been treated quite differently in the literature.

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Virtual Smart Metering in District Heating Networks via Heterogeneous Spatial-Temporal Graph Neural Networks

DGX agent

arXiv:2604.10166v1 Announce Type: cross Abstract: Intelligent operation of thermal energy networks aims to improve energy efficiency, reliability, and operational flexibility through data-driven contr

model-releasesarxiv-cs-ai
14 Apr 2026
Research

VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites

DGX agent

arXiv:2506.14629v3 Announce Type: replace-cross Abstract: Mosquito-borne diseases pose a major global health risk, requiring early detection and proactive control of breeding sites to prevent outbreak

researcharxiv-cs-cl
14 Apr 2026
Research

Visual Late Chunking: An Empirical Study of Contextual Chunking for Efficient Visual Document Retrieval

DGX agent

arXiv:2604.10167v1 Announce Type: cross Abstract: Multi-vector models dominate Visual Document Retrieval (VDR) due to their fine-grained matching capabilities, but their high storage and computational

researcharxiv-cs-cl
14 Apr 2026
← Previous
1…484485486487488…1371
Next →