AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,191 results
Model Releases

Optimization-Embedded Active Multi-Fidelity Surrogate Learning for Multi-Condition Airfoil Shape Optimization

DGX agent

arXiv:2603.17057v2 Announce Type: replace-cross Abstract: Active multi-fidelity surrogate modeling is developed for multi-condition airfoil shape optimization to reduce high-fidelity CFD cost while re

model-releasesarxiv-cs-lg
9 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Physics-Audited Agentic Discovery in Scientific Machine Learning

DGX agent

arXiv:2607.07379v1 Announce Type: new Abstract: In agentic scientific machine learning (SciML), large language model (LLM) agents can discover surrogate models and select one by an automated score, ty

agentsarxiv-cs-ai
9 Jul 2026
Agents

Power and Limitations of Aggregation in Compound AI Systems

DGX agent

arXiv:2602.21556v2 Announce Type: replace Abstract: When designing compound AI systems, a common approach is to query multiple copies of the same model and aggregate the responses to produce a synthes

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

Reinforcement Federated Learning Method Based on Adaptive OPTICS Clustering

DGX agent

arXiv:2306.12859v3 Announce Type: replace Abstract: Federated learning is a distributed machine learning technology, which realizes the balance between data privacy protection and data sharing computi

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Specification Grounding Drives Test Effectiveness for LLM Code

DGX agent

arXiv:2607.06636v1 Announce Type: cross Abstract: Large language models frequently generate code that appears correct on typical inputs yet fails on edge cases, invalid inputs, and other specification

model-releasesarxiv-cs-ai
9 Jul 2026
Safety

The Power of Backdoor Absorption in Community Training

DGX agent

arXiv:2607.06643v1 Announce Type: cross Abstract: Backdoor attacks severely threaten large-scale AI models. When model owners delegate training to external compute providers within a decentralized tra

safetyarxiv-cs-lg
9 Jul 2026
Model Releases

TimEE: End-to-end Time Series Classification via In-Context Learning

DGX agent

arXiv:2607.07500v1 Announce Type: cross Abstract: Time series classification (TSC) is dominated by a two-stage paradigm: train a feature encoder -- either from scratch on the target dataset or via pre

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

TRACE-Seg3D: Counterfactual Context Auditing For Robust 3D Glioma Segmentation Under Institutional Shift

DGX agent

arXiv:2607.07038v1 Announce Type: new Abstract: Medical image segmentation models can achieve strong benchmark performance while remaining sensitive to scanner, protocol, and institutional variation.

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

Tree-of-Thoughts Reasoning for Text-to-Image In-Context Learning

DGX agent

arXiv:2607.07117v1 Announce Type: cross Abstract: In text-to-image in-context learning (T2I-ICL), a model has to infer a latent compositional pattern from fewshot demonstrations for generating a query

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Unraveling Machine Behavior by Multi-Level Bias Analysis and Detection: Methodology and Application to Computer Vision

DGX agent

arXiv:2607.07236v1 Announce Type: new Abstract: This study investigates the presence and propagation of bias within Neural Networks through a comprehensive multi-level analysis spanning the learned la

model-releasesarxiv-cs-cv
9 Jul 2026
Model Releases

Auto-DSM Under the Lens: A Black-Box Evaluation Framework for LLM-Based DSM Generation

DGX agent

arXiv:2607.05985v1 Announce Type: new Abstract: This paper presents a black-box evaluation framework to systematically assess the ability of Large Language Models (LLMs) to generate Design Structure M

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Benchmarking the Robustness of Autonomous Driving to Environmental Illusions: A Lane Perception Perspective

DGX agent

arXiv:2607.05783v1 Announce Type: new Abstract: Environmental illusions (eg., shadows, reflections, and tire marks) are naturally existing yet overlooked phenomena in real-world driving environments.

model-releasesarxiv-cs-cv
8 Jul 2026
Applications

CCBENCH: Assessing LLM Cultural Competence via Implicitly Signaled Norms using Health Queries

DGX agent

arXiv:2607.05405v1 Announce Type: cross Abstract: To interact with users fairly and without stereotyping, AI models must display cultural competency, i.e., the ability to infer and adapt to a user's i

applicationsarxiv-cs-ai
8 Jul 2026
Model Releases

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade

DGX agent

arXiv:2607.06503v1 Announce Type: new Abstract: Large language model (LLM) agents solving multi-step tasks frequently commit to trajectories that are doomed to fail, yet continue to consume substantia

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

MAME: Multidimensional Adaptive Metamer Exploration with Human Perceptual Feedback

DGX agent

arXiv:2503.13212v3 Announce Type: replace Abstract: Alignment between human brain networks and artificial models has become an active research area in vision science and machine learning. A widely ado

safetyarxiv-cs-lg
8 Jul 2026
Local Ai

Measuring the practice of shared-decision making (OPTION12): An Investigation into Open-sourced Smaller LLMs (OS-sLLMs) for Better Privacy and Sustainability

DGX agent

arXiv:2607.06127v1 Announce Type: new Abstract: We present LLM4SDM, the first study of open-source smaller language models (OS-sLLMs) for automated assessment of shared decision making (SDM) using the

local-aiarxiv-cs-cl
8 Jul 2026
Model Releases

MobileWan: Closing the Quality Gap for Mobile Video Diffusion

DGX agent

arXiv:2607.06173v1 Announce Type: new Abstract: Recent advances in video diffusion have been driven by scaling transformer-based architectures to billions of parameters, substantially improving visual

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges

DGX agent

arXiv:2607.05904v1 Announce Type: new Abstract: Training a language model against its own reference-free judgments (the premise of self-rewarding, self-play, and LLM-as-a-judge pipelines) assumes a mo

model-releasesarxiv-cs-lg
8 Jul 2026
Model Releases

Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic SpeechLLMs

DGX agent

arXiv:2601.12494v3 Announce Type: replace-cross Abstract: Audio large language models (LLMs) enable unified speech understanding and generation, but adapting them to linguistically complex and dialect

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

PatchOptic for Shared-State LLM Workflows with Projected Views and Verified Structured Updates

DGX agent

arXiv:2607.05483v1 Announce Type: cross Abstract: Agentic workflows often operate over shared, structured state. Because LLM context windows are limited, each model invocation is typically shown only

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

PIPBench: A Profile-Inclusive Framework for Personalized Image Generation Evaluation

DGX agent

arXiv:2607.06440v1 Announce Type: new Abstract: Recent text-to-image models such as DALLE-3 excel at following diverse prompts yet remain blind to individual aesthetic preferences. We study personaliz

model-releasesarxiv-cs-cv
8 Jul 2026
Model Releases

PolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents

DGX agent

arXiv:2607.06008v1 Announce Type: new Abstract: Large language model (LLM) agents have shown strong performance in long-horizon tasks that require planning, tool use, and interaction with external env

model-releasesarxiv-cs-ai
8 Jul 2026
Research

Prompting Complexity: Shortest Prompts for Texts and Behaviors in LLMs

DGX agent

arXiv:2607.06145v1 Announce Type: new Abstract: In this paper, we define the quantity of prompting complexity: for a fixed instruction-tuned language model, what is the shortest plausible prompt that

researcharxiv-cs-cl
8 Jul 2026
Local Ai

SAMPLe: SAM-based Optimizer for Prompt Learning in VLMs

DGX agent

arXiv:2607.05727v1 Announce Type: new Abstract: Pre-trained Vision-Language Models (VLMs) like CLIP have proven highly effective as foundation models for various downstream applications. However, prom

local-aiarxiv-cs-cv
8 Jul 2026
Model Releases

Amortising Bayesian Experimental Design for Sequential Information Gathering in LLMs

DGX agent

arXiv:2607.03426v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong reasoning and world-knowledge capabilities, yet often struggle to gather information effectively across th

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Anchored Self-Play for Code Repair

DGX agent

arXiv:2607.03523v1 Announce Type: cross Abstract: Code repair is an important capability for language models (LMs): given a buggy program and unit tests, an LM must produce a fixed program that passes

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

ARCQuant: Boosting NVFP4 Quantization with Augmented Residual Channels for LLMs

DGX agent

arXiv:2601.07475v2 Announce Type: replace-cross Abstract: The emergence of fine-grained numerical formats like NVFP4 presents new opportunities for efficient Large Language Model (LLM) inference. Howe

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Auto: The AGI Compiler

DGX agent

arXiv:2607.04542v1 Announce Type: cross Abstract: Every LLM agent run re-derives its behavior token by token on a frontier model: brilliant, expensive, slow, and unbounded. We present Auto, a compiler

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Back to Basics: Improving Molecular Understanding in LLMs via SMILES-Graph Translation

DGX agent

arXiv:2607.03007v1 Announce Type: cross Abstract: Recent advances in molecular large language models have led to strong performance on molecular understanding and generation tasks, yet these gains oft

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Beyond Modality Fusion: Deep Ensembles for Multimodal Classification

DGX agent

arXiv:2607.05019v1 Announce Type: cross Abstract: In multimodal classification, late-fusion approaches classify concatenated modality-specific features extracted by unimodal neural networks. When moda

model-releasesarxiv-cs-cv
7 Jul 2026
Applications

Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning

DGX agent

arXiv:2607.04577v1 Announce Type: new Abstract: Code models strictly prioritize functional correctness, leaving software energy efficiency as an unoptimized byproduct. Training models to generate ener

applicationsarxiv-cs-lg
7 Jul 2026
Model Releases

CausalGame: Benchmarking Causal Thinking of LLM Agents in Games

DGX agent

arXiv:2607.04293v1 Announce Type: cross Abstract: Building AI Scientist agents with Large Language Models (LLMs) has recently attracted growing attention. Since scientific discovery fundamentally reli

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

CineMobile: On-Device Image-to-Video Diffusion for Cinematic Camera Motion Generation

DGX agent

arXiv:2607.03803v1 Announce Type: cross Abstract: The growing demand for image-to-video creation on mobile devices has increasingly focused on cinematic motion effects like bullet time, dolly zoom, sl

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Code Benchmarks Should Prioritize Rigor, Reliability, and Reproducibility

DGX agent

arXiv:2501.10711v5 Announce Type: replace-cross Abstract: Code-related benchmarks play a critical role in evaluating large language models (LLMs), yet their quality fundamentally shapes how the commun

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs

DGX agent

arXiv:2508.10031v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) have shown significant advancements in performance, various jailbreak attacks have posed growing safety and

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

CritiqueDriveVLM: From Verifier-Guided Reinforcement Learning to Latent Thought Distillation for Autonomous Driving

DGX agent

arXiv:2607.04179v1 Announce Type: cross Abstract: End-to-end Vision-Language Models (VLMs) show immense potential in autonomous driving. However, standard Supervised Fine-Tuning (SFT) often suffers fr

model-releasesarxiv-cs-ai
7 Jul 2026
Hardware

Data Driven Optimization of GPU efficiency for Distributed LLM-Adapter Serving

DGX agent

arXiv:2602.24044v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) adapters enable low-cost model specialization, but introduce complex caching and scheduling challenges in distribut

hardwarearxiv-cs-ai
7 Jul 2026
Model Releases

Discrete distributions are learnable from metastable samples

DGX agent

arXiv:2410.13800v4 Announce Type: replace-cross Abstract: Physically motivated stochastic dynamics are widely used to sample from high-dimensional distributions. However, such samplers often get trapp

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation

DGX agent

arXiv:2607.05147v1 Announce Type: new Abstract: Speculative decoding accelerates Large Language Model (LLM) inference by decoupling draft generation from target verification. While recent parallel dra

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Dual-Adaptive SAM3: Hierarchical Routing over Low-Rank Expert Layers for Parameter-Efficient Medical Image Segmentation

DGX agent

arXiv:2607.02571v1 Announce Type: new Abstract: The Segment Anything Model with Concepts (SAM3) heralds a new paradigm for open-vocabulary segmentation through natural language interaction, offering s

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

FUSE: FK-Steered Multi-Modal Flow Matching for Efficient Simulation-Based Posterior Estimation

DGX agent

arXiv:2607.05252v1 Announce Type: new Abstract: Simulation-Based Inference (SBI) is critical for scientific discovery, with generative models offering a promising path toward efficient inference. Howe

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

FuseMamba-VD: Dual Branch VideoMamba with Gated Class Token Fusion for Violence Detection

DGX agent

arXiv:2506.03162v3 Announce Type: replace-cross Abstract: The rapid proliferation of surveillance cameras has increased the demand for automated violence detection. While CNNs and Transformers have sh

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Governed MCP: Kernel-Level Tool Governance for AI Agents via Logit-Based Safety Primitives

DGX agent

arXiv:2604.16870v2 Announce Type: replace-cross Abstract: AI agents increasingly call external tools (file system, network, APIs) through the Model Context Protocol (MCP). These tool calls are the age

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Grokking Is Conditional and Fragile: A Fully-Tractable, Multi-Seed Study at 12K Parameters

DGX agent

arXiv:2607.05104v1 Announce Type: cross Abstract: Grokking -- the delayed onset of generalization long after a network has fit its training set - -is usually studied in models too large to read comple

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

GuideMe: Multi-Domain Task Guidance and Intervention in Streaming Video

DGX agent

arXiv:2607.02991v1 Announce Type: new Abstract: While multimodal Large Language Models (MLLMs) excel at offline video understanding, an interesting question of how far they are from serving as a real-

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

IRIS: An Intelligent Vision-Language System for Ocular Surface Diseases via Topic Tree and Scene-Driven VQA Generation

DGX agent

arXiv:2607.04344v1 Announce Type: cross Abstract: While Large Vision-Language Models (VLMs) demonstrate remarkable generic capabilities, their clinical reasoning in specialized domains like ocular sur

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

K9-Bench: Evaluating Multimodal LLMs on Canine-Centric Videos

DGX agent

arXiv:2607.02680v1 Announce Type: cross Abstract: MLLMs have shown strong zero-shot capabilities across diverse inputs such as across images, video, audio, and text. A crucial, yet underexplored, appl

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Learning When to Attend: Conditional Memory Access for Long-Context LLMs

DGX agent

arXiv:2603.17484v2 Announce Type: replace Abstract: Language models struggle to generalize beyond pretraining context lengths, limiting long-horizon reasoning and retrieval. Continued pretraining on l

model-releasesarxiv-cs-cl
7 Jul 2026
← Previous
1…347348349350351…1067
Next →