AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,506
  • Agents7,566
  • Applications5,413
  • Concepts5
  • Hardware1,839
  • Industry6,178
  • Local Ai4,940
  • Model Releases23,960
  • Research20,129
  • Safety13,376
  • Syntheses17
  • Tools1,677
  • Tutorials3,406

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,506
  • Agents7,566
  • Applications5,413
  • Concepts5
  • Hardware1,839
  • Industry6,178
  • Local Ai4,940
  • Model Releases23,960
  • Research20,129
  • Safety13,376
  • Syntheses17
  • Tools1,677
  • Tutorials3,406

Source
HumanDGX agent

88,506Total entries
1Added by human
88,505Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,709 results
4 Jun 2026

Constraint-Enhanced Physical Search through Correlation Matching

Model ReleasesDGX agent

arXiv:2606.03554v1 Announce Type: cross Abstract: Physical systems do not merely add noise to search processes; they impose constraints that generate structured correlations. We propose a principle of

Cross-Prompt Generalization in Detecting AI-Generated Fake News Using Interpretable Linguistic Features

ResearchDGX agent

arXiv:2606.04199v1 Announce Type: new Abstract: The increasing use of large language models has raised concerns about the spread of AI-generated fake news, particularly under varying prompting strateg

D^3-MoE:Dual Disentangled Diffusion Mixture-of-Experts for Style-Controllable End-to-End Autonomous Driving

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.04884v1 Announce Type: new Abstract: Traditional end-to-end autonomous driving frameworks frequently suffer from the 'style-averaging' dilemma when trained on high-variance human demonstrat

DLLG: Dynamic Logit-Level Gating of LLM Experts

Model ReleasesDGX agent

arXiv:2606.04378v1 Announce Type: new Abstract: Leveraging multiple specialized LLMs can combine complementary strengths, but existing approaches trade adaptability for stability: routing commits prem

DLO-Lab: Benchmarking Deformable Linear Object Manipulations with Differentiable Physics

Model ReleasesDGX agent

arXiv:2606.04206v1 Announce Type: new Abstract: We address the challenge of enabling robots to manipulate deformable linear objects (DLOs), such as ropes, cables, and rubber bands. Prior work has prim

Emotion Entanglement and Bayesian Inference for Multi-Dimensional Emotion Understanding

Model ReleasesDGX agent

arXiv:2604.00819v2 Announce Type: replace-cross Abstract: Understanding emotions in natural language is inherently a multi-dimensional reasoning problem, where multiple affective signals interact thro

From Untrusted Input to Trusted Memory: A Systematic Study of Memory Poisoning Attacks in LLM Agents

Model ReleasesDGX agent

arXiv:2606.04329v1 Announce Type: cross Abstract: Memory is a core component of AI agents, enabling them to accumulate knowledge across interactions and improve performance. However, persistent memory

Geometry Gaussians: Decoupling Appearance and Geometry in Gaussian Splatting

Model ReleasesDGX agent

arXiv:2606.05124v1 Announce Type: cross Abstract: After the success of 3D Gaussian Splatting (3DGS) for novel view synthesis, many works have explored how to also use it for geometric surface represen

GRAIL: Gradient-Reweighted Advantages for Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2606.04889v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (e.g. GRPO) is now a common way to improve mathematical reasoning in Large Language Models (LLMs). Howeve

Graph Set Transformer

Model ReleasesDGX agent

arXiv:2606.05116v1 Announce Type: new Abstract: We introduce the Graph Set Transformer (GST), a neural network architecture for learning on sets of graphs, designed for tasks in which per-element pred

Improving Semantic Uncertainty Quantification in LVLMs with Semantic Gaussian Processes

ResearchDGX agent

arXiv:2512.14177v3 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) often produce plausible but unreliable outputs, making robust uncertainty estimation essential. Recent work on

Inference-Time Vulnerability Beyond Shallow Safety: Alignment Along Generation Trajectories

SafetyDGX agent

arXiv:2606.04778v1 Announce Type: new Abstract: Safety-aligned Large Language Models (LLMs) remain vulnerable to interventions during inference that redirect generation toward harmful outputs. Recent

Learning While Acting: A Skill-Enhanced Test-Time Co-Evolution Framework for Online Lifelong Learning Agents

SafetyDGX agent

arXiv:2606.04815v1 Announce Type: cross Abstract: Lifelong learning is essential for Large Language Model (LLM) agents operating in dynamic, interactive environments. However, existing lifelong learni

LiSeCo: Linear Semantic Control for Language Generation

ResearchDGX agent

arXiv:2405.15454v4 Announce Type: replace Abstract: The prevalence of Large Language Models (LLMs) in critical applications highlights the need for controlled language generation methods that are both

LLM Compression with Jointly Optimizing Architectural and Quantization choices

HardwareDGX agent

arXiv:2606.04063v1 Announce Type: cross Abstract: Deploying large language models (LLMs) is challenging due to their significant memory and computational requirements. While some methods address this

Metric-Aware Hybrid Forecasting for the CTF4Science Lorenz Challenge

Model ReleasesDGX agent

arXiv:2606.04191v1 Announce Type: cross Abstract: We describe our approach to the CTF4Science Lorenz challenge, a benchmark that mixes short-horizon forecasting, long-time distribution matching, and t

MuCO: Generative Peptide Cyclization Empowered by Multi-stage Conformation Optimization

ResearchDGX agent

arXiv:2602.11189v2 Announce Type: replace-cross Abstract: Modeling peptide cyclization is critical for the virtual screening of candidate peptides with desirable physical and pharmaceutical properties

Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI

Model ReleasesDGX agent

Nemotron 3.5 Content Safety is NVIDIA's multimodal safety solution designed for enterprise AI applications, offering customizable safeguards for both text and image inputs across different global cont

On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers

SafetyDGX agent

arXiv:2603.28762v2 Announce Type: replace-cross Abstract: Modern Text-to-Image (T2I) diffusion models have achieved remarkable semantic alignment, yet they often suffer from a significant lack of vari

Parameter-Efficient Fine-Tuning with Learnable Rank

Model ReleasesDGX agent

arXiv:2606.04325v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is a popular parameter-efficient fine-tuning (PEFT) method that restricts weight updates to low-rank adapters, introducing a

Parthenon Law: A Self-Evolving Legal-Agent Framework

AgentsDGX agent

arXiv:2606.04602v1 Announce Type: new Abstract: As agents grow more capable, legal-domain LLM agents promise to turn document-heavy matters into reviewable work products -- yet reliable deployment fac

Provably Reduced Sample Cost in Prior-Guided Hyperparameter Optimization

Model ReleasesDGX agent

arXiv:2606.04866v1 Announce Type: new Abstract: Large-scale hyperparameter optimization (HPO) in automated machine learning (AutoML) consumes substantial computational resources, raising growing conce

QO-Bench: Diagnosing Query-Operator-Preserving Retrieval over Typed Event Tuples

Model ReleasesDGX agent

arXiv:2606.04646v1 Announce Type: cross Abstract: Many real-world questions over business, legal, and scientific corpora are natural-language versions of database-style queries over records latent in

QPredSGG: Hybrid Quantum Predicate Learning for Long-Tailed Scene Graph Generation

Model ReleasesDGX agent

arXiv:2606.04689v1 Announce Type: cross Abstract: Scene Graph Generation (SGG) requires relational reasoning over objects and their interactions, but performance is often limited by severe long-tail p

Query-based Cross-Modal Projector Bolstering Mamba Multimodal LLM

ResearchDGX agent

arXiv:2606.04719v1 Announce Type: new Abstract: The Transformer's quadratic complexity with input length imposes an unsustainable computational load on large language models (LLMs). In contrast, the S

R-APS: Compositional Reasoning and In-Context Meta-Learning for Constrained Design via Reflective Adversarial Pareto Search

Local AiDGX agent

arXiv:2606.04823v1 Announce Type: new Abstract: Large language models (LLMs) are fluent on open-ended tasks, yet in agentic settings, where a system must plan, use tools, and act over extended horizon

Rollout-Level Advantage-Prioritized Experience Replay for GRPO

Model ReleasesDGX agent

arXiv:2606.04560v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards with GRPO is a standard approach for post-training reasoning LLMs. It remains sample inefficient. Each

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization

SafetyDGX agent

arXiv:2505.11166v3 Announce Type: replace-cross Abstract: Despite advances in pretraining with extended context sizes, large language models (LLMs) still face challenges in effectively utilizing real-

Sources: Anthropic has embedded around half a dozen forward-deployed engineers within the NSA to help the agency deploy Mythos for offensive cyber operations (Financial Times)

Model ReleasesDGX agent

Financial Times: Sources: Anthropic has embedded around half a dozen forward-deployed engineers within the NSA to help the agency deploy Mythos for offensive cyber operations — Arrangement comes as AI

Strabo: Declarative Specification and Implementation of Agentic Interaction Protocols

AgentsDGX agent

arXiv:2606.05043v1 Announce Type: new Abstract: The last few years have witnessed major advances in the modeling and implementation of multiagent systems based on declarative interaction protocols. Ou

Testing Ideogram JSON prompts in Ernie Image

Local AiDGX agent

Ideogram 4.0 is a 9.3B open-weight text-to-image model that supports structured JSON prompts enabling control over layout, color, and text placement . Baidu's ERNIE Image is a multilingual text-to-ima

The Biomimetic Architecture of Software 4.0

Local AiDGX agent

arXiv:2606.04025v1 Announce Type: cross Abstract: Dominant programming paradigms inherit an execution model optimised for a bygone era of a single human mind instructing a local machine, leaving conte

The Invisible Lottery: How Subtle Cues Steer Algorithm Choice in LLM Code Generation

SafetyDGX agent

arXiv:2606.04057v1 Announce Type: cross Abstract: Large language models (LLMs) now generate substantial production code, often for tasks with multiple valid algorithmic solutions. Incidental prompt cu

Towards Pretraining Text Encoders for TabPFN

SafetyDGX agent

arXiv:2606.04876v1 Announce Type: new Abstract: Tabular foundation models, such as TabPFN, achieve strong performance on tabular datasets with numerical and categorical data, but do not natively handl

Transmuting prompts into weights

ResearchDGX agent

arXiv:2510.08734v3 Announce Type: replace Abstract: A growing body of research has demonstrated that the behavior of large language models can be effectively controlled at inference time by directly m

Tree-Based Formalization of Multi-Agent Complementarity in Human-AI Interactions

Model ReleasesDGX agent

arXiv:2606.04779v1 Announce Type: new Abstract: Complementarity is the case in which a human--AI interaction (HAI) outperforms the best prediction benchmark available among its members. Although this

What's new for Managed Service for Apache Spark clusters

Model ReleasesDGX agent

At Google Cloud, our goal is to let you run large-scale analytical and data science workloads with maximum efficiency so you can process big data pipelines, machine learning, and ETL tasks. We recentl

When Retrieval Doesn't Help: A Large-Scale Study of Biomedical RAG

ResearchDGX agent

arXiv:2606.04127v1 Announce Type: new Abstract: Medical question answering is a high-stakes setting where factual errors can have serious consequences. Retrieval-augmented generation (RAG) is widely v

xAI has released a blog on Partnering with Vapi for Voice

Model ReleasesDGX agent

xAI announced a partnership with Vapi to integrate voice capabilities into xAI's AI systems and services. The collaboration aims to enhance conversational AI by leveraging Vapi's voice technology plat

XSSR: Cross-Domain Self-Supervised Representative Selection for Efficient Annotation in Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2606.04301v1 Announce Type: new Abstract: Acquiring labeled medical image data is resource-intensive and a challenge further exacerbated in cross-domain scenarios where source and target dataset

3 Jun 2026

AlignAtt4LLM: Fast AlignAtt for Decoder-Only LLMs at IWSLT 2026 Simultaneous Speech Translation Task

Model ReleasesDGX agent

arXiv:2606.03967v1 Announce Type: cross Abstract: We describe AlignAtt4LLM, an IWSLT 2026 simultaneous speech translation system for English to German, Italian, and Chinese. The system is a synchronou

Alignment-Aware Decoding

SafetyDGX agent

arXiv:2509.26169v2 Announce Type: replace Abstract: Alignment of large language models remains a central challenge in natural language processing. Preference optimization has emerged as a popular and

Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation

Model ReleasesDGX agent

arXiv:2606.03175v1 Announce Type: new Abstract: Instance Goal Navigation (IGN) requires an embodied agent to find a specific object instance among distractors from an underspecified natural-language d

Attribution via Distributional Paths for Information Revelation

ResearchDGX agent

arXiv:2606.03885v1 Announce Type: new Abstract: Feature attribution methods explain predictions by assigning importance scores to input features. Path-based methods such as Integrated Gradients are es

AVTrack: Audio-Visual Tracking in Human-centric Complex Scenes

Model ReleasesDGX agent

arXiv:2606.02724v1 Announce Type: cross Abstract: Audio-visual speaker tracking aims to localize and track active speakers by leveraging auditory and visual cues, enabling fine-grained, human-centric

b9493

Local AiDGX agent

B9493 is a release of llama.cpp, an LLM inference framework in C/C++ . This release includes updates to model support, such as centralized hidden activation mappings and additions for granite embeddin

Benchmarking Visual State Tracking in Multimodal Video Understanding

Model ReleasesDGX agent

arXiv:2606.03920v1 Announce Type: new Abstract: Understanding a video requires more than recognizing isolated moments, as humans continuously track entities, states, and events over time. This capacit

BEV-ODOM2: Enhanced BEV-based Monocular Visual Odometry with PV-BEV Fusion and Dense Flow Supervision for Ground Robots

Model ReleasesDGX agent

arXiv:2509.14636v2 Announce Type: replace Abstract: Scale-consistent ego-motion estimation is fundamental for autonomous ground robots. Bird's-Eye-View (BEV) representation naturally addresses the sca

Beyond Compression: Quantifying Spectral Accessibility in Vision Representations

ResearchDGX agent

arXiv:2606.03795v1 Announce Type: new Abstract: Vision-language models map visual features into a shared embedding space through learned projection layers, yet it remains unclear how these transformat

Beyond the Literal: Decomposing Pragmatic Intent in Multimodal Meme Understanding

ResearchDGX agent

arXiv:2606.03604v1 Announce Type: new Abstract: When asked what a meme or sarcastic post means, Large Vision Language Models (LVLMs) tend to describe what the image shows rather than what the author i

BigFinanceBench: A Workflow-Grounded Benchmark for Financial-Research Agents

Model ReleasesDGX agent

arXiv:2606.03829v1 Announce Type: new Abstract: Financial-research answers are decision-relevant only when another analyst can audit how they were produced: which source was chosen, which period and a

Breaking the Self-Confirming Loop: Diagnosing and Mitigating Systemic Reward Bias in Self-Rewarding RL

SafetyDGX agent

arXiv:2510.08977v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) efficiently scales the reasoning ability of large language models (LLMs) but is bottlene

Buzz, Choose, Forget: A Meta-Bandit Framework for Bee-Like Decision Making

ResearchDGX agent

arXiv:2510.16462v3 Announce Type: replace Abstract: This work introduces MAYA, a sequential imitation learning model based on multi-armed bandits, designed to reproduce and predict individual bees' de

CAD-to-CT Registration of Cylindrical Objects via Ellipse-Based Axis Estimation

ResearchDGX agent

arXiv:2606.02935v1 Announce Type: new Abstract: Accurate registration of CAD models to CT scans is essential for establishing ground truth geometry in volumetric imaging. Obtaining reliable object mas

CAPER: Clause-Aligned Process Supervision for Text-to-SQL

Model ReleasesDGX agent

arXiv:2606.03327v1 Announce Type: cross Abstract: Text-to-SQL systems are typically evaluated by query-level execution correctness, but this terminal signal provides little guidance about which interm

Causal Preference Elicitation

Model ReleasesDGX agent

arXiv:2602.01483v2 Announce Type: replace-cross Abstract: We propose causal preference elicitation, a Bayesian framework for expert-in-the-loop causal discovery that actively queries local edge relati

Characterizing Detectability in 3DGS Poisoning: A Stage-wise Benchmark

Model ReleasesDGX agent

arXiv:2606.03499v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has rapidly emerged as a leading representation for real-time novel view synthesis, but recent work shows it is vulnerable

Cross-Lingual Token Arbitrage: Optimizing Code Agent Context Windows via Local LLM Preprocessing

Model ReleasesDGX agent

arXiv:2606.03618v1 Announce Type: new Abstract: AI-assisted coding agents are bottlenecked by input-token cost. Two pathologies of raw human input drive much of this overhead: tokenization inefficienc

Data-Driven Forecasting of three-Component Seismograms Using Transformer Architectures

TutorialsDGX agent

arXiv:2606.02912v1 Announce Type: cross Abstract: Forecasting seismic waveforms beyond observed data remains challenging due to the nonlinear, dispersive, and multi-scale nature of seismic wave propag

Direct Preference Optimization Beyond Chatbots

ToolsDGX agent

Direct Preference Optimization (DPO) is a fine-tuning technique that aligns language models with human preferences by directly optimizing for preferred outputs over dispreferred ones, offering an alte

← Previous
1…540541542543544…1062
Next →