AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
Human
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
5 Aug 2026

Interpreting Black-Box Large Language Models with Sentence-Level Energy Landscapes

Local AiDGX agent

arXiv:2608.02879v1 Announce Type: new Abstract: The widespread adoption of proprietary Large Language Models (LLMs) accessed strictly through closed APIs has created a critical challenge for responsib

Intertemporal Preference Steering in Qwen3 via Contrastive Activation Addition

Model ReleasesDGX agent

arXiv:2608.03892v1 Announce Type: new Abstract: We study linear representations of temporal horizon in the large language model Qwen3-32B and use them to change the model's time-related preferences, r

Inverted Detection and Control in Steering Vectors

Model ReleasesDGX agent

arXiv:2608.02957v1 Announce Type: new Abstract: Steering vectors (SVs) are widely used to influence the expression of concepts (e.g., truthfulness) in large language model outputs. A key assumption un

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structural Pruning

Model ReleasesDGX agent

arXiv:2507.14171v3 Announce Type: replace-cross Abstract: Importance-based structured pruning overwhelmingly relies on filter magnitude. This proxy is fundamentally flawed: due to scale invariance, fu

IR2Solve: Structured Intermediate Representations for Cost-Efficient Optimization Autoformulation

AgentsDGX agent

arXiv:2608.02641v1 Announce Type: cross Abstract: Large language models (LLMs) can translate natural-language optimization problems into solver-ready formulations, but direct code generation is brittl

IRIS: Visual-Semantic Binding for Forgery-Resistant Watermarking of Diffusion Images

ResearchDGX agent

arXiv:2608.03539v1 Announce Type: new Abstract: Most in-generation diffusion watermarks embed patterns independent of the image that carries them, and attackers transplant the marks onto images the ge

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details

AgentsDGX agent

arXiv:2608.03644v1 Announce Type: new Abstract: AI agents deployed in real-world settings must be capable of coordinating with humans and other AI agents they have not encountered before. Zero-shot co

ISEE: Interactive Semantic Enrichment for Database Fields

ApplicationsDGX agent

arXiv:2608.02604v1 Announce Type: new Abstract: LLM-based agents are increasingly being deployed for data-related tasks, including data sense-making, exploration, and retrieval. However, their perform

Joint Affine Spectral Shaping: Coupling Weight and Bias Updates Beyond Weight-Only Muon

SafetyDGX agent

arXiv:2608.02991v1 Announce Type: new Abstract: Matrix spectral optimizers reshape weight-update spectra but usually delegate vector-valued biases to a separate optimizer. We study whether this separa

JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion

Model ReleasesDGX agent

arXiv:2608.03974v1 Announce Type: new Abstract: Real-time video editing requires low-latency causal generation with bounded computational resources while preserving source fidelity and long-term tempo

JudgeArena: A Unified Framework for Reproducible LLM-Judge Evaluation

Model ReleasesDGX agent

arXiv:2608.02620v1 Announce Type: new Abstract: LLM-as-a-judge evaluation has become a dominant paradigm for ranking language models, yet the ecosystem remains fragmented: most benchmarks ship their o

Keep the Needle, Prune the Haystack: Defect-Preserving Token Pruning for Efficient Zero-Shot Anomaly Detection

Local AiDGX agent

arXiv:2608.03681v1 Announce Type: new Abstract: Zero-shot visual anomaly detection has achieved remarkable progress, with recent vision-only approaches further improving performance while simplifying

Kernel weighted importance sampling for off-policy evaluation in contextual bandits

SafetyDGX agent

arXiv:2607.15067v2 Announce Type: replace Abstract: This article presents a novel estimator for performing off-policy evaluation using only offline data for contextual bandits. The proposed estimator,

KernelBrain: Coarse-to-Fine, Budget-Aware Search for Agentic GPU Kernel Optimization

SafetyDGX agent

arXiv:2608.02611v1 Announce Type: cross Abstract: Automating GPU kernel optimization remains difficult in practice: generated variants can violate correctness constraints, runtime measurements are noi

KnowHal: A Knowledge-Driven Benchmark for Comprehensive Multimodal Hallucination Evaluation

Model ReleasesDGX agent

arXiv:2608.03782v1 Announce Type: new Abstract: Hallucination remains a critical challenge for developing trustworthy Multimodal Large Language Models (MLLMs). While existing benchmarks mainly focus o

Knowing the Form, Not the Function: Automatically Auditing Answer--Authority Decoupling in Legal Benchmarks

Model ReleasesDGX agent

arXiv:2608.02621v1 Announce Type: cross Abstract: Legal benchmarks typically score final answers even when models also state legal authority. We test whether answer correctness can serve as a proxy fo

LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension

AgentsDGX agent

arXiv:2608.02915v1 Announce Type: cross Abstract: Domain-specific Instruction Set Architecture eXtensions (ISAX) are widely adopted in the RISC-V ecosystem to accelerate emerging workloads, but implem

LAEF: A Lead-Agnostic ECG Foundation Model Towards Point-of-Care Diagnostics

Model ReleasesDGX agent

arXiv:2608.03690v1 Announce Type: new Abstract: Point-of-care cardiac devices such as smartwatches and handheld ECG recorders typically capture 1--2 leads, yet existing ECG foundation models are archi

Language Models Encode the Contextual Truth of Propositions

ResearchDGX agent

arXiv:2608.03035v1 Announce Type: new Abstract: Prior work has shown that LLMs encode the truth of factual propositions along linear directions in activation space. It's unclear how these representati

Language-Specialized Multi-Teacher On-Policy Distillation for Multilingual LLM-Based ASR

SafetyDGX agent

arXiv:2608.03610v1 Announce Type: new Abstract: Modern LLM-based ASR systems have established multilingual capability as a standard feature, leveraging large-scale multilingual corpora and LLMs' cross

Large language models for partial differential equation workflows

ApplicationsDGX agent

arXiv:2608.03600v1 Announce Type: new Abstract: Partial differential equations (PDEs) become actionable in science and engineering not as isolated formulae, but as executable workflows that connect mo

Large Language Models provide support for the parallelogram theory of analogy

Local AiDGX agent

arXiv:2603.19066v2 Announce Type: replace-cross Abstract: Four-term word analogies (A:B::C:D) are classically modeled geometrically as parallelograms: adding the vector B-A+C produces D. Recent work s

Latent Reward Registers for Diffusion Preference Alignment

Model ReleasesDGX agent

arXiv:2608.03929v1 Announce Type: cross Abstract: Aligning diffusion models with human preferences usually relies on a sparse terminal reward evaluated on the final generated samples, presenting a sev

LatentGuard: Efficient and Inspectable Latent Reasoning for LLM Safeguards

SafetyDGX agent

arXiv:2608.03838v1 Announce Type: new Abstract: Reasoning-based guard models improve LLM safeguards, but decoding explicit rationales for every interaction makes them costly to deploy. Although latent

LDU-Bench: Multimodal LLM Evaluation for Lithography Defect Understanding under Layout-Varying Circuit Backgrounds

Model ReleasesDGX agent

arXiv:2608.03078v1 Announce Type: new Abstract: Multimodal large language models have demonstrated strong defect recognition capability in industrial anomaly detection. However, in lithography review,

LeanMem: Simple and Efficient Long-Term Memory for LLM Agents

Model ReleasesDGX agent

arXiv:2608.03463v1 Announce Type: new Abstract: Long-term memory is essential for LLM-based agents to sustain interactions and reliably leverage distant history. However, existing memory systems typic

Learning a Vector-Symbolic Model for Socio-Cultural Tasks

ResearchDGX agent

arXiv:2608.02807v1 Announce Type: cross Abstract: How can we better represent the impact of sociocultural structures on decision making in computational cognitive models? Modeling this impact requires

Learning and Clustering on Temporal Graphs: Principles, Primitives, and Pooling

HardwareDGX agent

arXiv:2608.03696v1 Announce Type: new Abstract: This work focuses on the problem of learning on temporal graphs, with particular emphasis on the task of clustering: obtaining coarse-grained representa

Learning Attribute-aware Representations for Few-shot Scene Text Segmentation

SafetyDGX agent

arXiv:2504.11164v2 Announce Type: replace Abstract: Supervised scene text segmentation has achieved notable progress in recent years. However, its development is largely constrained by the scarcity of

Learning Biomechanically Plausible Human Motion from Sparse Radar Point Clouds

ResearchDGX agent

arXiv:2608.03637v1 Announce Type: new Abstract: Radar-based human pose estimation has focused on improving learning algorithms while representing the body as unconstrained keypoint coordinates. We add

Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents

SafetyDGX agent

arXiv:2608.03606v1 Announce Type: new Abstract: Clinical development is sequential decision-making under uncertainty, where a sponsor must plan a portfolio of experiments from heterogeneous evidence.

Learning Context-Aware Motion Priors for Humanoid Control

SafetyDGX agent

arXiv:2608.03234v1 Announce Type: new Abstract: Motion priors provide powerful guidance for learning naturalistic humanoid behaviors. However, existing methods typically learn a general, task-agnostic

Learning Molecular Representations from Cellular Phenotypes with Structure Preservation

SafetyDGX agent

arXiv:2608.02688v1 Announce Type: cross Abstract: Phenotypic drug discovery enables the discovery of functional relationships between molecular structures and cellular responses. However, existing mul

Learning Music Style for Piano Arrangement Through Cross-Modal Bootstrapping

SafetyDGX agent

arXiv:2608.03050v1 Announce Type: cross Abstract: What is music style? Though often described using text labels such as 'swing,' 'classical,' or 'emotional,' the real style remains implicit and hidden

Less Traffic, Better Outcomes: Competition-Aware Request Dispatch in Real-Time Ad Exchanges

SafetyDGX agent

arXiv:2608.03705v1 Announce Type: new Abstract: Real-time bidding (RTB) ad exchanges typically forward nearly all incoming requests to demand-side platforms (DSPs), even though only a small fraction r

Leveraging System-Level Observations to Inform Bayesian Learning of Model Parameters for Quantitative Verification

ApplicationsDGX agent

arXiv:2608.03489v1 Announce Type: cross Abstract: Combining Bayesian learning and quantitative verification is a powerful toolset for analysing key quantitative properties of software systems, like re

Light-Loco-Parkour: Versatile Perceptive Whole-Body Locomotion via Multi-Skill Distillation

SafetyDGX agent

arXiv:2608.02653v1 Announce Type: new Abstract: Existing humanoid whole-body control systems still fall short of the way humans move through cluttered terrain: they either track expressive whole-body

Lightweight 3D Object Detection via Mamba-Based Knowledge Distillation

SafetyDGX agent

arXiv:2608.03490v1 Announce Type: cross Abstract: 3D object detection using light detection and ranging (LiDAR) sensors requires a balance between accuracy and computational efficiency for onboard per

Lightweight Chunk Selection for Mobile Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2608.03148v1 Announce Type: cross Abstract: RAG improves the factual grounding of LLM by incorporating external knowledge, but deploying RAG on mobile and edge devices remains challenging becaus

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation

HardwareDGX agent

arXiv:2608.03701v1 Announce Type: cross Abstract: World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticip

LiteMVS: Efficient Multi-View Stereo with Foundation Distillation and Expert Aggregation

ResearchDGX agent

arXiv:2608.03851v1 Announce Type: new Abstract: Real-time 3D perception is crucial for robotics, augmented reality, and embodied intelligence applications. Existing multi-view stereo (MVS) methods pri

LiveEvalBench: Toward Open-World Evaluation for Web Generation

AgentsDGX agent

arXiv:2608.03689v1 Announce Type: new Abstract: Large language models are increasingly capable of synthesizing executable frontend projects, yet existing benchmarks still treat web generation as a sta

LLaDA MoE v2: Scaling Mixture-of-Experts Diffusion Language Models

ResearchDGX agent

arXiv:2608.03457v1 Announce Type: new Abstract: Diffusion language models (dLLMs) offer an alternative to autoregressive (AR) language modeling, yet the scaling behavior of Mixture-of-Experts (MoE) dL

LLM-Derived Priors for Thompson Sampling in Cold-Start Comment Recommendation

SafetyDGX agent

arXiv:2608.03382v1 Announce Type: cross Abstract: Multi-armed bandit algorithms, especially Thompson sampling, are widely used in online recommendation. Despite their ability to adapt from online feed

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs

HardwareDGX agent

arXiv:2608.03036v1 Announce Type: cross Abstract: Large Language Models (LLMs) are integrated into software systems and AI services, making efficient LLM serving a concern for software engineering. Se

LLMs Can Annotate Attribution Graphs

ResearchDGX agent

arXiv:2608.02632v1 Announce Type: new Abstract: Circuit tracing is an exciting technique for revealing the internal computation of language models, but it requires a time-intensive manual step of grou

LoBoost: Fast Model-Native Local Conformal Prediction for Gradient-Boosted Trees

Local AiDGX agent

arXiv:2602.22432v2 Announce Type: replace-cross Abstract: Gradient-boosted decision trees are among the strongest off-the-shelf predictors for tabular regression, but point predictions alone do not qu

LoCA: Forward-Only LLM Tuning after One-Shot Calibration with Local Credit Assignment

Model ReleasesDGX agent

arXiv:2608.03020v1 Announce Type: new Abstract: Parameter-efficient post-training reduces the number of trainable parameters, but still requires repeated end-to-end backpropagation through the frozen

Localize, Don't Beautify: Client-Side Control of Image-Editing APIs for Cosmetic Surgery Previews

Model ReleasesDGX agent

arXiv:2608.02841v1 Announce Type: new Abstract: Ask a commercial image editor to preview a cosmetic procedure and it will often change more of the face than the request names: a nose edit can also smo

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images

Model ReleasesDGX agent

arXiv:2608.03322v1 Announce Type: new Abstract: Medical visual grounding connects free-form clinical queries to spatial evidence in medical images and is an important component of interpretable medica

Logic Before Language: Pre-pretraining on Formal Derivations Fosters Skill Acquisition and Compressibility

ResearchDGX agent

arXiv:2608.03930v1 Announce Type: cross Abstract: Pre-pretraining language models (LMs) on symbolic data can accelerate and improve natural language acquisition. However, existing pre-pretraining task

LogitScope: A Framework for Analyzing LLM Uncertainty Through Information Metrics

ApplicationsDGX agent

arXiv:2603.24929v2 Announce Type: replace Abstract: Understanding and quantifying uncertainty in large language model (LLM) outputs is critical for reliable deployment. However, traditional evaluation

Long-term Traffic Scene Prediction via Polynomial Representations in Autonomous Driving

SafetyDGX agent

arXiv:2608.03330v1 Announce Type: new Abstract: This thesis addresses fundamental challenges in traffic scene prediction for autonomous driving by introducing robust and computationally efficient mode

Looking under the Wrong Lamppost: On the Limitations of Automated Translation Quality Estimation

Model ReleasesDGX agent

arXiv:2608.03577v1 Announce Type: new Abstract: Automation of Translation Quality Estimation (QE) has emerged as a widely discussed approach to managing translation quality at scale, and a growing num

LoopMTP: A looped transformer guided by latent multi-token prediction

Model ReleasesDGX agent

arXiv:2608.03624v1 Announce Type: new Abstract: Looped transformers have emerged as a parameter-efficient alternative to scaling depth for strong reasoning. By reusing one stack of layers across T ite

Low-Dimensional High-Leverage Subspace Optimization: Beyond Full-Parameter Coupled Training for Neural Network Quantization

Model ReleasesDGX agent

arXiv:2608.03919v1 Announce Type: new Abstract: Low-bit quantization suffers severe accuracy degradation on compact networks, rooted in the dominant full-parameter coupled training paradigm that ignor

M-GATE: Multilingual Grammar, Accuracy in Translation, and Efficiency Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2608.03803v1 Announce Type: new Abstract: Multilingual language models are deployed across a hundred or more languages, yet most benchmarks test whether a model can perform a task _in_ a languag

MAFIA: Query-Only Memory Attacks via Probing and Factual Injection against Audited LLM Agents

AgentsDGX agent

arXiv:2608.03844v1 Announce Type: new Abstract: Memory-augmented LLM agents rely on rich context for long-horizon reasoning and acting, yet their memory modules expose a persistent attack surface for

Maglev: Sliding Recurrent Memory

Model ReleasesDGX agent

arXiv:2608.02870v1 Announce Type: new Abstract: We introduce ours{}, a recurrent Transformer architecture with fixed-size memory that generalizes sliding-window attention while remaining parallelizabl

MambaTS: Improved Selective State Space Models for Long-term Time Series Forecasting

SafetyDGX agent

arXiv:2405.16440v2 Announce Type: replace-cross Abstract: In recent years, Transformers have become the de-facto architecture for long-term time series forecasting (LTSF), yet they face challenges ass

← Previous
1…8182838485…998
Next →