AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,587 results
Model Releases

Diagnosing and Mitigating Retrieval Bottlenecks in LLM-Based Cold-Start Recommendation

DGX agent

arXiv:2606.29947v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as rerankers in recommender systems, with the expectation that semantic understanding will help in

model-releasesarxiv-cs-lg
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DialogPII: A multilingual dataset of synthetic dialog transcripts to detect personal information

DGX agent

arXiv:2606.30312v1 Announce Type: new Abstract: Conversational data collected in domains such as healthcare or social sciences is a valuable resource for research and automated analysis. However, resp

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Does Verbose Chain-of-Thought Really Help? In-Distribution Evidence that Content, Not Length, Matters

DGX agent

arXiv:2606.30128v1 Announce Type: new Abstract: Chain-of-thought (CoT) prompting improves LLM reasoning, but the source is contested: do the intermediate steps help because they carry useful semantic

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Early Warning Signals for OpenVLA Failure under Visual Distribution Shift

DGX agent

arXiv:2606.29699v1 Announce Type: cross Abstract: Vision Language Action models combine perception, language grounding, and control in a single policy, but their failures are hard to diagnose once vis

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Efficient Unlearning with Privacy Guarantees

DGX agent

arXiv:2507.04771v2 Announce Type: replace-cross Abstract: Privacy protection laws, such as the GDPR, grant individuals the right to request the forgetting of their personal data not only from database

researcharxiv-cs-lg
30 Jun 2026
Model Releases

Ensemble Learning Based Classification Algorithm Recommendation

DGX agent

arXiv:2101.05993v2 Announce Type: replace-cross Abstract: Selecting an appropriate classification algorithm for a given data set remains a challenging problem in data mining and machine learning. Exis

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

DGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Expert Evaluation of Clinical AI Tools on Real Point-of-Care Clinical Queries

DGX agent

arXiv:2606.28960v1 Announce Type: new Abstract: Physicians now pose millions of clinical questions to AI tools each week, yet these tools are evaluated largely on hypothetical or exam-style questions,

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Forensic Trajectory Signatures for Agent Memory Poisoning Detection

DGX agent

arXiv:2606.30566v1 Announce Type: cross Abstract: We discover a behavioral invariant in LLM agents under persistent memory poisoning: in architectures where routing information is retrieved through ob

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Geometrically Principled Randomized Optimization for Efficient LLM Training

DGX agent

arXiv:2510.01878v2 Announce Type: replace Abstract: Low-rank gradient optimization for large language models is currently divided into two categories: structured methods that rigorously identify subsp

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Governance Decay: How Context Compaction Silently Erases Safety Constraints in Long-Horizon LLM Agents

DGX agent

arXiv:2606.22528v2 Announce Type: replace Abstract: Modern LLM agents increasingly rely on context compaction, summarization, or eviction to keep long-running sessions within a token budget. We show t

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Improved Predictive Performance and Interpretability for Mesomorphic Neural Networks Using Local Fidelity Regularization

DGX agent

arXiv:2606.29951v1 Announce Type: new Abstract: Interpretable Mesomorphic Neural Networks (IMNs) offer a promising framework that combines the predictive power of deep neural networks with the interpr

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

LEDGER: Scaling Agentic Document Editing with Dependency-aware Graph Retrieval

DGX agent

arXiv:2606.28379v1 Announce Type: cross Abstract: We introduce LEDGER to tackle the novel context engineering challenge of agentic document editing, where localized edits to long, structured documents

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive Dashboard

DGX agent

arXiv:2606.30005v1 Announce Type: new Abstract: Long-horizon tool agents are bottlenecked by how their context grows toward the limits of the context window. Recent systems make context management age

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

LoRAShield: Data-Free Editing Alignment for Secure Personalized LoRA Sharing

DGX agent

arXiv:2507.07056v2 Announce Type: replace-cross Abstract: The proliferation of Low-Rank Adaptation (LoRA) models has democratized personalized text-to-image generation, enabling users to share lightwe

safetyarxiv-cs-lg
30 Jun 2026
Model Releases

Making Multimodal LLMs Reliable Chart Data Extractors: A Benchmark and Training Framework

DGX agent

arXiv:2606.29808v1 Announce Type: cross Abstract: Chart data extraction, which reverse-engineers data tables from chart images, is essential for reproducibility, analysis, retrieval, and redesign. Exi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

meta-pipe: An LLM-agent pipeline for end-to-end automated systematic review and meta-analysis

DGX agent

arXiv:2606.28363v1 Announce Type: cross Abstract: Objective: To describe the architecture and design rationale of meta-pipe, an open-source large language model (LLM)-agent pipeline that integrates th

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Mitigating Batch Effects in Histopathology via Language-Mediated Robust Embedding Generation

DGX agent

arXiv:2606.28697v1 Announce Type: cross Abstract: Pathology foundation models (PFMs) have demonstrated strong potential across clinical and scientific applications, yet their performance is often hind

researcharxiv-cs-cl
30 Jun 2026
Agents

Monte Carlo Query Search: Active Capability Assessment of AI Agents

DGX agent

arXiv:2512.16733v3 Announce Type: replace Abstract: Black-box AI (BBAI) systems, including foundation-model agents, are increasingly used for sequential decision making. Safe deployment requires metho

agentsarxiv-cs-ai
30 Jun 2026
Safety

muFlow: Leveraging Average Images for Improving Generalisation of Deepfake Faces Detectors

DGX agent

arXiv:2606.30528v1 Announce Type: new Abstract: Current generative models, including GANs and diffusion models, have reached an outstanding level of photorealism, posing significant risks to privacy a

safetyarxiv-cs-cv
30 Jun 2026
Safety

Multimodal Representation Alignment for Cross-modal Information Retrieval

DGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs

DGX agent

arXiv:2606.30026v1 Announce Type: cross Abstract: Audiovisual arts encompass diverse creative disciplines, including cinema, visual arts, stage performance, and game design, where artistic meaning ari

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Never Skip a Batch: Dense Learning of Temporal GNNs via Adaptive Pseudo-Supervision

DGX agent

arXiv:2505.12526v2 Announce Type: replace Abstract: Temporal graph networks suffer from irregular supervision in realworld dynamic graphs, as most minibatches contain few labeled events. The lack of l

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

OmniCoT: A Benchmark for Global and Multi-Step Panoramic Reasoning

DGX agent

arXiv:2606.30378v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated promising spatial reasoning capabilities, while these abilities remain underexplored in the e

model-releasesarxiv-cs-cv
30 Jun 2026
Research

Orca: The World is in Your Mind

DGX agent

arXiv:2606.30534v1 Announce Type: new Abstract: We introduce Orca, an initial instantiation of a general world foundation model. Orca learns a unified world latent space from multimodal world signals

researcharxiv-cs-cv
30 Jun 2026
Model Releases

PGE-SAM: Prompt-Guided Feature Enhancement for Interactive Segmentation under Degradation

DGX agent

arXiv:2606.30477v1 Announce Type: new Abstract: Segment Anything Model (SAM) has revolutionized promptable image segmentation with strong zero-shot generalization. However, its performance degrades su

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

DGX agent

arXiv:2602.11351v2 Announce Type: replace Abstract: Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

RiverONE: Generating Knowledge-Intensive VLM by Simulated Quantum Machines

DGX agent

arXiv:2606.29966v1 Announce Type: cross Abstract: Quantum computing provides a powerful paradigm for representing and transforming high-dimensional information through superposition, entanglement, and

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

DGX agent

arXiv:2606.29894v1 Announce Type: cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theore

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Scalable Synthesis of distributed LLM workloads through Symbolic Tensor Graphs

DGX agent

arXiv:2511.10480v3 Announce Type: replace-cross Abstract: Optimizing the performance of large language models (LLMs) on large-scale AI training and inference systems requires a scalable and expressive

researcharxiv-cs-ai
30 Jun 2026
Model Releases

SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents

DGX agent

arXiv:2603.29139v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled agentic systems to translate natural-language intent into executable scientific visuali

model-releasesarxiv-cs-ai
30 Jun 2026
Tutorials

See, Think, Learn: A Self-Taught Multimodal Reasoner

DGX agent

arXiv:2512.02456v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in integrating visual perception with language understanding. However, effecti

tutorialsarxiv-cs-cl
30 Jun 2026
Local Ai

Smooth Scaling Laws Hide Stepwise Token Learning

DGX agent

arXiv:2606.29858v1 Announce Type: new Abstract: Language model loss follows remarkably regular scaling laws over model and data size, yet it remains unclear why the aggregate loss should exhibit a pow

local-aiarxiv-cs-cl
30 Jun 2026
Local Ai

Spectral Perturbation of the Empirical Fisher Information Matrix under Weight Quantization

DGX agent

arXiv:2606.28432v1 Announce Type: cross Abstract: We study the spectral perturbation of the empirical Fisher Information Matrix (FIM) of a parametric statistical model under two structured perturbatio

local-aiarxiv-cs-ai
30 Jun 2026
Model Releases

StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley

DGX agent

arXiv:2507.07445v3 Announce Type: replace Abstract: Autonomous agents navigating human society must master both production activities and social interactions, yet existing benchmarks rarely evaluate t

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents

DGX agent

arXiv:2606.28480v1 Announce Type: cross Abstract: As large language models and harness frameworks continue to advance, agents operating in terminals are increasingly capable of performing a broader ra

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Variance Reduction on the Camera Axis: Multi-View Score Distillation for 3D

DGX agent

arXiv:2606.29964v1 Announce Type: new Abstract: Score distillation turns a pretrained 2D diffusion model into a 3D generator, but the per-step gradient is estimated from a single randomly chosen view:

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

When Does Overlap Help? OSU-Mem and a Cell-Conditional Analysis of Trajectory Memory for LLM Agents

DGX agent

arXiv:2606.28376v1 Announce Type: cross Abstract: Long-horizon large language model (LLM) agents accumulate interaction trajectories that quickly exceed any practical prompt budget, and existing memor

model-releasesarxiv-cs-ai
30 Jun 2026
Research

When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets

DGX agent

arXiv:2606.29248v1 Announce Type: new Abstract: Vegetable prices in Sri Lanka are highly volatile because the market is largely import-isolated, so supply disruptions quickly drive prices up. This stu

researcharxiv-cs-lg
30 Jun 2026
Safety

Words Speak Louder Than Code: Investigating Cognitive Heuristics in LLM-Based Code Vulnerability Detection

DGX agent

arXiv:2606.30587v1 Announce Type: cross Abstract: Researchers and practitioners increasingly apply Large Language Models (LLMs) for automated vulnerability detection. Recent work has shown that LLMs a

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

XRAG: eXamining the Core -- Benchmarking Foundational Components in Advanced Retrieval-Augmented Generation

DGX agent

arXiv:2412.15529v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) synergizes the retrieval of pertinent data with the generative capabilities of Large Language Models (LLM

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

A Multi-Attribute Latent Space for Visual Analysis of Watches

DGX agent

arXiv:2606.27897v1 Announce Type: new Abstract: We present a design rationale, embedding model, and interactive visual-analysis system for exploring large wristwatch collections through heterogeneous

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

A Tree-of-Thoughts Inspired Hybrid Approach for Legal Case Judgement Summarization using LLMs

DGX agent

arXiv:2606.28044v1 Announce Type: new Abstract: In recent times, Large Language Models (LLMs) are increasingly being used for legal case judgement summarization. Most prior works have tried traditiona

model-releasesarxiv-cs-cl
29 Jun 2026
Model Releases

Building a Scalable, Reproducible, Evaluatable, and Closed-Loop Simulation Environment Foundation for Embodied Intelligence Cloud-Native Simulation Infrastructure for Embodied Intelligence Training, Evaluation, and Data Collection

DGX agent

arXiv:2606.27962v1 Announce Type: new Abstract: This paper presents a cloud-native simulation infrastructure framework for embodied intelligence that supports large-scale training, standardized evalua

model-releasesarxiv-cs-ro
29 Jun 2026
Safety

Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training

DGX agent

arXiv:2602.16065v2 Announce Type: replace-cross Abstract: As artificial intelligence (AI)-generated content proliferates, models are increasingly trained on their own outputs, risking progressive degr

safetyarxiv-cs-ai
29 Jun 2026
Tutorials

CBD: API-Only LLM Black-Box Unlearning through Controlled Behavioral Divergence

DGX agent

arXiv:2606.27683v1 Announce Type: cross Abstract: Edge devices increasingly invoke large language models (LLMs) through API services for context aware edge intelligence, while edge generated data may

tutorialsarxiv-cs-ai
29 Jun 2026
Model Releases

DMind Benchmark: Toward a Holistic Assessment of LLM Capabilities across the Web3 Domain

DGX agent

arXiv:2504.16116v4 Announce Type: replace-cross Abstract: The Web3 ecosystem, underpinned by cryptographic primitives and decentralized consensus, represents a high-stakes environment where software v

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

FoggyTrust: Robust Federated Learning with Hierarchical Trust Networks

DGX agent

arXiv:2606.27622v1 Announce Type: new Abstract: Byzantine-robust federated learning seeks to protect distributed model training from malicious or corrupted clients without requiring access to their pr

model-releasesarxiv-cs-lg
29 Jun 2026
← Previous
1…406407408409410…1075
Next →