AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Local Ai

A Robust and Explainable Transformer-Based Framework for Phishing Email Detection

DGX agent

arXiv:2511.12085v3 Announce Type: replace-cross Abstract: Phishing and related cyber threats are becoming increasingly sophisticated, with email-based phishing remaining the most persistent attack vec

local-aiarxiv-cs-ai
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Acceptance-Test-Driven Evaluation Protocols for Business-Centric LLM Systems

DGX agent

arXiv:2606.02755v1 Announce Type: cross Abstract: Large language model (LLM) applications are increasingly expected to satisfy deterministic institutional requirements while relying on probabilistic g

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

Agent libOS: A Library-OS-Inspired Runtime for Long-Running, Capability-Controlled LLM Agents

DGX agent

arXiv:2606.03895v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from request-response assistants into long-running software actors: they maintain state across model ca

local-aiarxiv-cs-ai
3 Jun 2026
Research

Bootstrap Your Generator: Unpaired Visual Editing with Flow Matching

DGX agent

arXiv:2606.03911v1 Announce Type: new Abstract: Modern generative models possess a deep understanding of visual content, yet training them for image editing typically requires massive datasets of pair

researcharxiv-cs-cv
3 Jun 2026
Model Releases

Compress then Merge: From Multiple LoRAs into One Low-Rank Adapter

DGX agent

arXiv:2606.03723v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) enables parameter-efficient specialization of foundation models, but the proliferation of task-specific adapters fragments ca

model-releasesarxiv-cs-lg
3 Jun 2026
Research

Decomposing how prompting steers behavior

DGX agent

arXiv:2606.03093v1 Announce Type: new Abstract: Prompting steers large language models (LLMs) and vision-language models (VLMs) without weight updates, but it remains unclear how instruction changes r

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Decoupled Smart Contract Audits: Lightweight LLM Framework via Distillation and Aggregation

DGX agent

arXiv:2606.03128v1 Announce Type: cross Abstract: Smart contracts face critical security challenges that require thorough auditing in decentralized web services. While Large Language Models (LLMs) hav

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Do Value Vectors in Deep Layers Need Context from the Residual Stream?

DGX agent

arXiv:2606.02780v1 Announce Type: new Abstract: The success of the transformer architecture as the backbone of modern LLMs is in large part due to its use of attention layers. An attention layer follo

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Don't Forget Your Embeddings: Robust Knowledge Erasure via Precise Editing of Embeddings

DGX agent

arXiv:2606.03695v1 Announce Type: new Abstract: As language models are increasingly deployed in real-world applications, the ability to erase specific knowledge from them becomes critical for safety a

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Dynamics of Cognitive Heterogeneity: Investigating Behavioral Biases in Multi-Stage Supply Chains with LLM-Based Simulation

DGX agent

arXiv:2604.17220v2 Announce Type: replace-cross Abstract: Modeling coordination among generative agents in complex multi-round decision-making presents a core challenge for AI and operations managemen

model-releasesarxiv-cs-ai
3 Jun 2026
Research

EntangleCodec: A Unified Discrete Audio Tokenizer via Semantic-Acoustic Entanglement

DGX agent

arXiv:2606.02739v1 Announce Type: cross Abstract: Audio tokenizers serve as the discrete interface between continuous audio and Audio Language Models (ALMs), but existing tokenizers often struggle to

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Experience-Driven Dynamic Exits for LLMs with Reinforcement Learning

DGX agent

arXiv:2606.03113v1 Announce Type: new Abstract: Large Language Models suffer from slow autoregressive inference. While self-speculative decoding accelerates this process, its efficiency is hampered by

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Forecasting Conceptual Diffusion in Science: The Case of Quantum Computing

DGX agent

arXiv:2606.03919v1 Announce Type: cross Abstract: Understanding and anticipating scientific change requires models that distinguish between endogenous consolidation and exogenous diffusion of scientif

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

From Prompt to Service: An SLM-Based Agent Orchestration Gateway for AI-Driven Virtual Worlds

DGX agent

arXiv:2606.03557v1 Announce Type: new Abstract: As generative AI capabilities expand, AI-driven virtual worlds face a growing architectural challenge. Users interact through in-world interfaces in mul

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

G^2C-MT: Graph-Guided Context Selection for Document-Level Machine Translation

DGX agent

arXiv:2606.03078v1 Announce Type: new Abstract: Effective document-level machine translation (DocMT) requires capturing long-range discourse dependencies. Recent work has explored retrieval-based and

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Identifying Quantum Structure in AI Language: Evidence for Evolutionary Convergence of Human and Artificial Cognition

DGX agent

arXiv:2511.21731v2 Announce Type: replace-cross Abstract: We present the results of cognitive tests on conceptual combinations, performed using specific Large Language Models (LLMs) as test subjects.

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

IdiomX A Multilingual Benchmark for Idiom Understanding, Retrieval, and Interpretation

DGX agent

arXiv:2606.02584v1 Announce Type: cross Abstract: Idiomatic expressions remain a persistent challenge for natural language processing because their meanings are often non-compositional, context-depend

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Leave it to the Specialist: Repair Sparse LLMs with Sparse Fine-Tuning via Sparsity Evolution

DGX agent

arXiv:2505.24037v3 Announce Type: replace Abstract: Sparse large language models (LLMs) offer an attractive direction toward efficient deployment, but adapting them to downstream tasks remains challen

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

MLSkip: Data Skipping for ML Filters via Lightweight Metadata

DGX agent

arXiv:2606.03946v1 Announce Type: cross Abstract: Database vendors recently released AI functions that can be used in filter predicates. As such functions often rely on costly, black-box ML models, th

model-releasesarxiv-cs-lg
3 Jun 2026
Research

PointAction: 3D Points as Universal Action Representations for Robot Control

DGX agent

arXiv:2606.03943v1 Announce Type: new Abstract: Video-Action Models (VAMs) leverage the broad visual dynamics captured by pre-trained video diffusion models, offering a promising path toward generaliz

researcharxiv-cs-ro
3 Jun 2026
Model Releases

PubTables-v2: A new large-scale dataset for full-page and multi-page table extraction

DGX agent

arXiv:2512.10888v3 Announce Type: replace Abstract: Table extraction (TE) is a key challenge in document understanding. Traditional approaches detect tables first, then recognize their structure. Rece

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation

DGX agent

arXiv:2606.03784v1 Announce Type: new Abstract: Embodied chain-of-thought (CoT) aims to bridge linguistic reasoning and robotic control, but its effective form and integration strategy remain underexp

model-releasesarxiv-cs-ro
3 Jun 2026
Research

Rex: A Family of Reversible Exponential (Stochastic) Runge-Kutta Solvers

DGX agent

arXiv:2502.08834v4 Announce Type: replace-cross Abstract: Deep generative models based on neural differential equations have become state-of-the-art for many generation tasks. These models rely on ODE

researcharxiv-cs-ai
3 Jun 2026
Model Releases

scTranslation: A Comprehensive Benchmark for Single-Cell Multi-Omics Modality Translation

DGX agent

arXiv:2606.03906v1 Announce Type: new Abstract: Simultaneous measurement of multiple omics modalities in single cells enables researchers to gain a more comprehensive understanding of cellular states

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

DGX agent

arXiv:2606.02911v1 Announce Type: new Abstract: Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in

safetyarxiv-cs-cl
3 Jun 2026
Model Releases

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning

DGX agent

arXiv:2606.03503v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have achieved remarkable progress thanks to Reinforcement Learning with Verifiable Rewards (RLVR) on Chain-of-Thoughts (Co

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Unified Video-Action Joint Denoising for Dexterous Action and Data Generation

DGX agent

arXiv:2606.03868v1 Announce Type: new Abstract: Recent world action models leverage video foundation models by aligning broad visual-dynamics priors with executable robot actions. We revisit this alig

safetyarxiv-cs-cv
3 Jun 2026
Model Releases

AblationBench: Evaluating Automated Planning of Ablations in Empirical AI Research

DGX agent

arXiv:2507.08038v3 Announce Type: replace-cross Abstract: Language model agents are increasingly used to automate scientific research, yet evaluating their scientific contributions remains a challenge

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

ArrythML: An Autoencoder-Based TinyML Approach for On-Device Arrhythmia Detection on Resource-Constrained Embedded Systems

DGX agent

arXiv:2606.02256v1 Announce Type: new Abstract: Our work presents a method for ECG segmentation and arrhythmia detection using Tiny Machine Learning (TinyML) models for real-time, on-device inference

local-aiarxiv-cs-lg
2 Jun 2026
Research

Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations

DGX agent

arXiv:2511.20295v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) are minimal and semantically meaningful modifications of the input of a model that alter the model predictions. T

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Before and After Temperature: A Distributional View of Creative LLM Generation

DGX agent

arXiv:2606.01451v1 Announce Type: new Abstract: Reference-free evaluation of large language model (LLM) creativity relies on perplexity, entropy, and top-1 margin. We show that a much stronger signal

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Benchmark Dataset for Catalysis on 2D MXenes

DGX agent

arXiv:2606.00794v1 Announce Type: cross Abstract: Merging first-principles calculations with machine learning (ML), we aim to accelerate the exploration of catalytic behaviour in novel materials. We f

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Benchmarking LLM-as-a-Judge for Long-Form Output Evaluation

DGX agent

arXiv:2606.01629v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly used for long-form generation, reliably evaluating long-form outputs has become a critical challenge. L

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Benchmarking Local LLMs for Natural-Language-to-SQL Querying in Biopharmaceutical Manufacturing: An Empirical Benchmark on Consumer-Grade Hardware

DGX agent

arXiv:2606.01338v1 Announce Type: new Abstract: Biopharmaceutical manufacturing organizations operate under regulatory frameworks such as FDA guidance, EU Good Manufacturing Practice (GMP), and the EU

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Beware of the Batch Size: Hyperparameter Bias in Evaluating LoRA

DGX agent

arXiv:2602.09492v2 Announce Type: replace-cross Abstract: Low-rank adaptation (LoRA) is a standard approach for fine-tuning large language models, yet its many variants report conflicting empirical ga

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification

DGX agent

arXiv:2510.24078v2 Announce Type: replace Abstract: Text-to-image (T2I) models are increasingly used for synthetic dataset generation, but generating effective synthetic training data for classificati

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Beyond Semantic Understanding: Preserving Collaborative Frequency Components in LLM-based Recommendation

DGX agent

arXiv:2508.10312v2 Announce Type: replace Abstract: Recommender systems in concert with Large Language Models (LLMs) present promising avenues for generating semantically-informed recommendations. How

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention

DGX agent

arXiv:2512.10414v2 Announce Type: replace Abstract: Recently, reinforcement learning (RL) has become a common choice in enhancing the reasoning capabilities of vision-language models (VLMs). Consideri

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Bridging the Sim-to-Real Gap in Semiconductor Visual Program Synthesis via Input Binarization

DGX agent

arXiv:2606.02434v1 Announce Type: new Abstract: Precise parametric control over circuit geometry is essential for semiconductor inspection, yet obtaining sufficient real training data remains costly.

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Child-directed speech facilitates production, not comprehension, in BabyLMs

DGX agent

arXiv:2606.01045v1 Announce Type: new Abstract: Recent studies suggest that child-directed speech is not conducive to language learning in BabyLMs. However, current evaluations focus predominantly on

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs

DGX agent

arXiv:2606.00898v1 Announce Type: new Abstract: Large language models systematically hallucinate legal citations -- fabricating statute references, citing repealed provisions, and confusing jurisdicti

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents

DGX agent

arXiv:2606.02568v1 Announce Type: new Abstract: Clinical practice is not the selection of an answer from enumerated options: a physician gathers heterogeneous information incrementally and commits to

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

ClinTutor-R1: Advancing Scalable and Robust One-to-Many Alignment in Clinical Socratic Education

DGX agent

arXiv:2512.05671v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have achieved remarkable success in dyadic (one-on-one) instruction, they face significant challenges in One-to-M

safetyarxiv-cs-cl
2 Jun 2026
Model Releases

Consistency evaluation of benchmarks used for causal discovery

DGX agent

arXiv:2606.01789v1 Announce Type: new Abstract: In graphical causal model, causal discovery aims to construct a causal graph based on numerical data and domain knowledge in plain text. However, the ev

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging

DGX agent

arXiv:2606.01717v1 Announce Type: new Abstract: Instruction tuning aligns large language models, including multimodal ones, with diverse user intents, but scaling to heterogeneous mixtures is hindered

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Disentanglement-Based Equivariant Learning for Compositional VQA

DGX agent

arXiv:2606.02168v1 Announce Type: new Abstract: Compositional visual question answering (VQA) represents a challenging yet fundamental task that requires models to comprehend novel combinations of pre

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking

DGX agent

arXiv:2606.01240v1 Announce Type: new Abstract: The demand for powerful instruction following and reasoning capability of large language models (LLMs) has promoted rapid development of retrieval-augme

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Escaping the BLEU Trap: A Signal-Grounded Framework with Decoupled Semantic Guidance for EEG-to-Text Decoding

DGX agent

arXiv:2603.03312v3 Announce Type: replace-cross Abstract: Decoding natural language from non-invasive EEG signals is a promising yet challenging task. However, current state-of-the-art models remain c

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…360361362363364…1074
Next →