AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Research

Linguistic Monoculture in LLM-Assisted Language Use

DGX agent

arXiv:2607.27134v1 Announce Type: cross Abstract: Writing and communication are increasingly mediated by large language models (LLMs) that are being used to draft, revise and polish text. Although suc

researcharxiv-cs-cl
30 Jul 2026
Model Releases

Position: Evaluation Scores Are Perishable Knowledge Claims

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.26191v1 Announce Type: cross Abstract: Evaluation methodologies for language models increasingly combine multiple signals, from automated metrics and LLM-as-judge ratings to human assessmen

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

DGX agent

arXiv:2607.26339v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) systems ground large language models (LLMs) in external corpora, but this reliance exposes them to corpus poisoning

model-releasesarxiv-cs-lg
30 Jul 2026
Safety

Self-Adaptive Learning and Model Predictive Control for Tracking Unknown Dynamics with No Regret

DGX agent

arXiv:2607.26370v1 Announce Type: cross Abstract: We propose a self-adaptive online learning for control method for tracking unknown target dynamics. The target dynamics can exhibit switching behavior

safetyarxiv-cs-lg
30 Jul 2026
Model Releases

The Reliability of LLMs for Medical Diagnosis: An Examination of Consistency, Manipulation, and Contextual Awareness

DGX agent

arXiv:2503.10647v2 Announce Type: replace Abstract: This study evaluated the diagnostic reliability of two Large Language Models (LLMs), Google Gemini 2.0 Flash and OpenAI ChatGPT-4o, across three dim

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Beyond Static Costs: Learning-Dynamics Aware Loss Functions for Long-Tailed Classification

DGX agent

arXiv:2607.25830v1 Announce Type: new Abstract: Deep learning models in computer vision face significant challenges when trained on long-tailed datasets, where a few majority classes dominate while ma

model-releasesarxiv-cs-cv
29 Jul 2026
Safety

Contrastive Weak-to-strong Generalization

DGX agent

arXiv:2510.07884v3 Announce Type: replace-cross Abstract: Weak-to-strong generalization provides a promising paradigm for scaling large language models (LLMs) by training stronger models on samples fr

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following

DGX agent

arXiv:2607.25398v1 Announce Type: new Abstract: Language-model agents are increasingly deployed under standing instructions: a system prompt, a policy file, or a skills document is placed in context,

model-releasesarxiv-cs-ai
29 Jul 2026
Research

KQFuzz: Knowledge-Guided Fuzzing for Quantum Libraries via Large Language Models

DGX agent

arXiv:2607.25647v1 Announce Type: cross Abstract: As quantum computing continually improves, ensuring the reliability and correctness of quantum libraries has become increasingly critical. To this end

researcharxiv-cs-ai
29 Jul 2026
Model Releases

LaP-Forensics: Latent-Pixel Consistency Guided Multimodal Reasoning for Deepfake Detection

DGX agent

arXiv:2607.25962v1 Announce Type: new Abstract: Recent generative models can produce images with few obvious visual artifacts, weakening detectors and explanations that rely only on surface appearance

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

Localized Adaptation Reveals Distinct Learning Signatures in Transformers

DGX agent

arXiv:2607.25663v1 Announce Type: new Abstract: Transformer adaptation is typically distributed across model depth, even when the intended change is narrow. We investigate how adaptation site shapes w

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

M^2PO: Multi-Perspective Multi-Pair Preference Optimization for Machine Translation

DGX agent

arXiv:2510.13434v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with human preferences is pivotal for Machine Translation (MT), yet current approaches are often hindered by m

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

Multimodal Hybrid Retrieval-Augmented Generation for Scientific Document Understanding using Open-Source SLMs

DGX agent

arXiv:2607.24799v1 Announce Type: cross Abstract: Large Language Models tend to hallucinate when answering domain-specific ques tions from scientific documents without prior fine-tuning. Currently, me

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Parallel Decoding Distillation for Fast Image and Video Generation

DGX agent

arXiv:2607.26004v1 Announce Type: new Abstract: Generation in video diffusion or flow models is computationally expensive due to the slow and iterative sampling process. Current state-of-the-art (SOTA

model-releasesarxiv-cs-cv
29 Jul 2026
Model Releases

PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

DGX agent

arXiv:2607.25485v1 Announce Type: new Abstract: Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Pri

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision

DGX agent

arXiv:2607.25321v1 Announce Type: new Abstract: Video diffusion models generate visually compelling content but routinely violate elementary physics when the subject involves fluids: liquid columns br

model-releasesarxiv-cs-ai
29 Jul 2026
Local Ai

Athena-Brain Technical Report: An Efficient Robot Brain for General Intelligence and Embodied Interaction

DGX agent

arXiv:2607.18985v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated remarkable capabilities in language understanding, reasoning, and world knowledge. As embodied agents

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

Benchmarking LLMs for Verilog Design Flows

DGX agent

arXiv:2607.22759v1 Announce Type: cross Abstract: Large language models (LLMs) show promise in code generation, but their capabilities to produce correct, synthesizable hardware description language (

model-releasesarxiv-cs-lg
28 Jul 2026
Safety

Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining

DGX agent

arXiv:2607.23175v1 Announce Type: cross Abstract: Reducing toxicity is often framed as a global alignment problem, yet perceptions of harmful language are subjective and context-dependent. We present

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Beyond ICA: Identifiability by Symmetry Breaking

DGX agent

arXiv:2607.23182v1 Announce Type: cross Abstract: We prove the identifiability of deep generative models (DGMs) with piecewise-affine (PWA) decoders and Gaussian mixture model (GMM) priors, in a purel

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

CausalGate: Causal Importance Distillation for Transformer Module Pruning

DGX agent

arXiv:2607.22720v1 Announce Type: cross Abstract: Existing adaptive inference methods for Large Language Models rely on observational heuristics, such as hidden-state similarity or activation magnitud

model-releasesarxiv-cs-cl
28 Jul 2026
Research

Differencing the Diffusion Trajectory toward Uncertain Components for Time Series Forecasting

DGX agent

arXiv:2607.22599v1 Announce Type: new Abstract: Diffusion models have become a widely used framework for probabilistic time series forecasting, modeling the distribution of future values given an obse

researcharxiv-cs-ai
28 Jul 2026
Model Releases

DuoAD: Leveraging [CLS] Dual Characteristics for Training-Free Few-Shot Anomaly Detection

DGX agent

arXiv:2607.23924v1 Announce Type: cross Abstract: Vision foundation models have enabled strong training-free anomaly detection (AD). However, most existing approaches rely primarily on independent loc

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

DynaCalKV: Key-Value Cache Compression via Head Grouping and Adaptive Rank Allocation

DGX agent

arXiv:2607.24331v1 Announce Type: new Abstract: As the inference phase of Large Language Models (LLMs) requires handling long context windows, the Key-Value (KV) cache initially appears to address thi

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios

DGX agent

arXiv:2607.23722v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as agents that interact with stateful environments over multiple steps: gathering hidden informat

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Evidence Attribution in Visual Document Understanding without Coordinates or Region Labels

DGX agent

arXiv:2607.24651v1 Announce Type: cross Abstract: Reliable visual document understanding requires a model to attribute each answer to the evidence regions that support it. Recent benchmarks and system

researcharxiv-cs-cl
28 Jul 2026
Model Releases

Harmonized Interpretable ECG Waveform Features for Robust Cross-Dataset Clinical Prediction

DGX agent

arXiv:2607.23412v1 Announce Type: new Abstract: Electrocardiograms (ECGs) are widely used for cardiovascular risk prediction, yet models often fail to transfer across hospitals because of protocol, po

model-releasesarxiv-cs-lg
28 Jul 2026
Research

INSIGHT: Spatially resolved survival modelling from routine histology crosslinked with molecular profiling reveals prognostic epithelial-immune axes in stage II/III colorectal cancer

DGX agent

arXiv:2512.22262v2 Announce Type: replace-cross Abstract: Routine histology contains rich prognostic information in stage II/III colorectal cancer, much of which is embedded in complex spatial tissue

researcharxiv-cs-lg
28 Jul 2026
Tutorials

Like a Baby: Visually Situated Neural Language Acquisition

DGX agent

arXiv:1805.11546v3 Announce Type: replace-cross Abstract: We examine the benefits of visual context in training neural language models to perform next-word prediction. A multi-modal neural architectur

tutorialsarxiv-cs-ai
28 Jul 2026
Research

Not Forgotten: Implementation and Evaluation of a Personalized Episodic Memory for the Humanoid Robot Head Kim

DGX agent

arXiv:2607.24190v1 Announce Type: cross Abstract: Social robots that rely on large language models for conversation are unable to retain information across sessions. This absence of memory violates so

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Nova3D: Code-Native Generation of Programmable 3D Assets

DGX agent

arXiv:2607.22738v1 Announce Type: cross Abstract: Current 3D generative models mostly produce a final surface: a visually strong but largely opaque mesh. Interactive 3D worlds need more than a surface

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

ObsDriveBench: Benchmarking Multimodal Understanding under Adverse Weather with Observability Awareness

DGX agent

arXiv:2607.23537v1 Announce Type: new Abstract: Autonomous driving under adverse weather remains a critical challenge, yet existing vision-language benchmarks mainly evaluate under standard conditions

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

PANOPTICON: A PII-Based Assemblage of Naturalistic Output Tokens for Investigating Privacy Leakage Within LLM Context Window

DGX agent

arXiv:2607.22695v1 Announce Type: new Abstract: Large Language Models (LLMs) are capable of generalizing human language for the completion of never-before-seen tasks, leading to widespread deployment.

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Scale Weight Decay and Train Better

DGX agent

arXiv:2607.23777v1 Announce Type: cross Abstract: The discovery of scaling laws has motivated training neural networks on ever increasing quantities of data. This is typically done with a constant dec

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Seesaw: Accelerating Training by Balancing Learning Rate and Batch Size Scheduling

DGX agent

arXiv:2510.14717v2 Announce Type: replace-cross Abstract: Increasing the batch size during training -- a ''batch ramp'' -- is a promising strategy to accelerate large language model pretraining. While

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning

DGX agent

arXiv:2607.22732v1 Announce Type: new Abstract: LLM-based game agents often perform poorly on more complex tasks. This work examines whether these failures are linked to limited spatial reasoning and

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

StanceFlip: A Comprehensive Multi-Dimensional Benchmark for Multimodal Conversational Stance Flipping Forecasting

DGX agent

arXiv:2607.24191v1 Announce Type: cross Abstract: Conversational stance detection has shifted from static text analysis to dynamic multimodal modeling. However, existing benchmarks exhibit three key l

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Tailored untruths: How personalisation challenges LLM safeguards

DGX agent

arXiv:2510.12993v3 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive disinformation, yet little is known about how effectively they personalise it across lan

safetyarxiv-cs-cl
28 Jul 2026
Safety

The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation

DGX agent

arXiv:2607.24720v1 Announce Type: cross Abstract: Multi-turn long-horizon planning is critical for foundation model agents, yet how to fundamentally improve it remains unclear. Existing models are tra

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Verbalized Particle Posterior: Bayesian Inference over Natural Language Hypotheses

DGX agent

arXiv:2607.22961v1 Announce Type: cross Abstract: Verbalized Machine Learning (VML) parameterizes a model as a natural-language prompt that an LLM evaluates as f(x; theta). The framework is interpreta

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

When LLM Defenses Backfire: Characterizing Safety, Performance, and Cost Trade-offs

DGX agent

arXiv:2607.24392v1 Announce Type: cross Abstract: Jailbreak defenses are essential for protecting large language models (LLMs), but they can also introduce secondary costs that weaken model utility. W

model-releasesarxiv-cs-lg
28 Jul 2026
Research

WISERouter: LLM Routing with Workload Budget Constraint

DGX agent

arXiv:2607.23765v1 Announce Type: cross Abstract: Large language models (LLMs) achieve impressive performance across multiple domains, but using the most capable model for every query is prohibitive a

researcharxiv-cs-ai
28 Jul 2026
Model Releases

DM3D: Dynamic Mamba via Offset-Guided Feature Resampling for Point Cloud Understanding

DGX agent

arXiv:2512.03424v4 Announce Type: replace Abstract: State Space Models (SSMs) model long token sequences of point cloud with linear complexity, but require an unordered point cloud to be serialized. E

model-releasesarxiv-cs-cv
27 Jul 2026
Model Releases

Language-Aware Distillation for Multilingual Instruction-Following Speech LLMs with ASR-Only Supervision

DGX agent

arXiv:2603.07025v2 Announce Type: replace Abstract: Speech Large Language Models (LLMs) that understand and follow instructions in many languages are useful for real-world interaction, but are difficu

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

LMEB: Long-horizon Memory Embedding Benchmark

DGX agent

arXiv:2603.12572v5 Announce Type: replace Abstract: Memory embeddings are crucial for memory-augmented systems, such as OpenClaw, but their evaluation is underexplored in current text embedding benchm

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

Offline Vision-Language Navigation with Geometric Goal Localization for Outdoor Environments

DGX agent

arXiv:2607.22226v1 Announce Type: new Abstract: Foundation-model-based vision-language navigation (VLN) has advanced autonomous robot navigation by enabling robots to interpret natural-language instru

model-releasesarxiv-cs-ro
27 Jul 2026
Model Releases

Optimization of time-consuming experimental conditions using pseudo-experimental data guided by adaptive polynomial regression

DGX agent

arXiv:2607.22238v1 Announce Type: new Abstract: Bayesian optimization (BO) is an optimization method that sequentially proposes the next candidate explainable variables for optimizing target variables

model-releasesarxiv-cs-lg
27 Jul 2026
Model Releases

RadSight: Towards Perceptually Reliable Multimodal Radiology Image Understanding

DGX agent

arXiv:2607.22293v1 Announce Type: new Abstract: Medical multimodal large language models (MLLMs) are increasingly expected to perform complex image understanding tasks, yet their reliability is often

model-releasesarxiv-cs-cv
27 Jul 2026
← Previous
1…277278279280281…1058
Next →