AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction

DGX agent

arXiv:2605.24657v1 Announce Type: new Abstract: Major LLM platforms deploy models in an inference-only configuration: the model serves requests but never updates per-user weights. Users must repeatedl

model-releasesarxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Beyond Literal Translation: Evaluating Cultural Effectiveness in Social Media UGC

DGX agent

arXiv:2605.25626v1 Announce Type: new Abstract: Social media platforms enable large-scale cross-lingual communication, but translating user-generated content (UGC) remains challenging due to its infor

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Beyond Query Memorization: Large Language Model Routing with Query Decomposition and Historical Matching

DGX agent

arXiv:2605.25558v1 Announce Type: new Abstract: Optimizing the trade-off among predictive performance and computational cost is a central focus in the deployment of Large Language Models (LLMs). Curre

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Beyond Summaries: Structure-Aware Labeling of Code Changes with Large Language Models

DGX agent

arXiv:2605.26100v1 Announce Type: cross Abstract: Code review is a critical practice in software engineering, yet the growing scale and frequency of code patches in modern projects, together with the

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

BODHI: Precise OS Kernel Specification Inference

DGX agent

arXiv:2605.23931v1 Announce Type: new Abstract: The formal verification of operating system kernels requires precise specifications that capture the intended behavior of system calls. Writing these sp

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Bridging On-Device and Cloud LLMs for Collaborative Reasoning: A Unified Methodology for Local Routing and Post-Training

DGX agent

arXiv:2509.24050v4 Announce Type: replace Abstract: Device-cloud collaboration holds promise for deploying large language models (LLMs), leveraging lightweight on-device models for efficiency while re

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Building an Adversarial Malware Dataset by Family and Type: Generation, Evasion, and Poisoning Evaluation

DGX agent

arXiv:2605.25937v1 Announce Type: cross Abstract: We present a dataset of adversarial malware samples derived from the public RawMal-TF collection of real-world malware binaries. Using a suite of adve

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Can LLMs Time Travel? Enhancing Temporal Consistency in Legal Agentic Search through Reinforcement Learning

DGX agent

arXiv:2605.25920v1 Announce Type: cross Abstract: While large language models (LLMs) augmented with agentic search capabilities show promise for legal reasoning, they overlook a fundamental constraint

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Can LoRA Fusion Support Cross-Domain Tasks in Cloud-Edge Collaboration?

DGX agent

arXiv:2605.23913v1 Announce Type: cross Abstract: Cloud-hosted large language models (LLMs) commonly rely on LoRA for domain adaptation, yet domain data are distributed across multiple edge devices an

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Cascade-KDE: Robust Time-Series Restoration under Out-of-Distribution Impulse Corruptions

DGX agent

arXiv:2605.24055v1 Announce Type: cross Abstract: Real-world time-series data in industrial sensing, healthcare, and energy systems is often corrupted by a mixture of Gaussian noise and occasional lar

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Causal Tongue-Tie: LLMs Can Encode Causal Direction, But Their Yes/No Outputs Fail to Express

DGX agent

arXiv:2605.25891v1 Announce Type: cross Abstract: We find a mismatch between what large language models encode about a causal question and what they answer. On anti-commonsense CLadder items, a fixed

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CausaLab: A Scalable Environment for Interactive Causal Discovery Toward AI Scientists

DGX agent

arXiv:2605.26029v1 Announce Type: new Abstract: We introduce CausaLab, a scalable environment for evaluating interactive causal discovery by LLM agents. Unlike prior evaluations, CausaLab evaluates bo

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Chain-of-Thought Hijacking

DGX agent

arXiv:2510.26418v4 Announce Type: replace Abstract: Large Reasoning Models (LRMs) improve task performance through extended inference-time reasoning. Although previous studies suggest that longer reas

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

ChainLearn: A Blockchain-Based Capacity-Aware Framework for Federated Ensemble Learning

DGX agent

arXiv:2605.24418v1 Announce Type: new Abstract: Federated learning is used in medical imaging where privacy prohibits centralizing data. Standard federated algorithms assume homogeneous hardware, iden

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

ChaosBench-Logic v2: Evaluating LLM Logical Reasoning over Dynamical Systems at Scale

DGX agent

arXiv:2605.24305v1 Announce Type: cross Abstract: Standard accuracy on binary reasoning benchmarks hides critical failure modes: prior collapse, inconsistency under paraphrase, and inability to reason

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference

DGX agent

arXiv:2510.02361v2 Announce Type: replace-cross Abstract: Transformer-based large models excel in natural language processing and computer vision, but face severe computational inefficiencies due to t

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CITYREP: A Unified Benchmark for Urban Representations Across Cities, Tasks, and Modalities

DGX agent

arXiv:2605.26036v1 Announce Type: new Abstract: Urban representation learning encodes complex urban environments into general-purpose embeddings for diverse downstream tasks and emerging urban foundat

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Clarification Is Not Enough: Post-Clarification Answering Remains the Bottleneck in Multi-Turn QA

DGX agent

arXiv:2605.25204v1 Announce Type: new Abstract: Pluralistic alignment requires systems to adapt to diverse user values, communication styles, and contextual assumptions. We believe that a foundational

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Claw-Anything: Benchmarking Always-On Personal Assistants with Broader Access to User's Digital World

DGX agent

arXiv:2605.26086v1 Announce Type: new Abstract: Large language model agents are increasingly envisioned as always-on personal assistants with access to anything relevant in the user's digital world. Y

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CMAP: Cross-Modal Adaptive Prompting for Multi-Domain Task-Incremental Learning

DGX agent

arXiv:2605.25708v1 Announce Type: cross Abstract: Multi-domain task-incremental learning requires a model to sequentially acquire knowledge across visually diverse domains without forgetting prior tas

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Code2UML: Agentic LLMs with context engineering for scalable software visualization

DGX agent

arXiv:2605.24453v1 Announce Type: cross Abstract: Large Language Model (LLM)-based code analysis tools are adopted to automate software documentation tasks. However, the scalability of these approache

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-Teacher On-Policy Distillation

DGX agent

arXiv:2605.25378v1 Announce Type: cross Abstract: Customized image editing aims to equip pre-trained diffusion models with specific visual effects using limited paired data, typically via Low-Rank Ada

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Committed SAE-Feature Traces for Audited-Session Substitution Detection in Hosted LLMs

DGX agent

arXiv:2604.18179v2 Announce Type: replace-cross Abstract: Hosted-LLM providers have a silent-substitution incentive: advertise a stronger model while serving cheaper replies. Probe-after-return scheme

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Complement Submodular Information Measures for Balanced and Robust Data Selection

DGX agent

arXiv:2605.24779v1 Announce Type: cross Abstract: Submodular optimization has become a fundamental paradigm for data selection, retrieval, summarization, and representation learning due to its ability

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models

DGX agent

arXiv:2605.25765v1 Announce Type: cross Abstract: Concept unlearning aims to erase a target concept from a pretrained text-to-image diffusion model without retraining. Closed-form methods are attracti

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Conformalised imprecise inference for robust extrapolation under limited data

DGX agent

arXiv:2605.25882v1 Announce Type: new Abstract: Recent advances in uncertainty quantification increasingly emphasise the distinction between aleatory and epistemic uncertainty in machine learning, mot

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Context-Instrumental Data Distillation for Kubernetes Manifest Generation: Method and Experimental Evaluation

DGX agent

arXiv:2605.25835v1 Announce Type: cross Abstract: This paper examines the specialization of Small Language Models (SLMs) with up to 4 billion parameters for generating artifacts in domain-specific lan

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions

DGX agent

arXiv:2605.24279v1 Announce Type: new Abstract: A frontier language model's acknowledged 'helpful programming assistant' persona does not survive long agentic-coding sessions in the deployment regime

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Continual Speaker Identity Unlearning with Minimal Interference

DGX agent

arXiv:2605.25962v1 Announce Type: cross Abstract: Machine unlearning removes designated concepts or knowledge from pre-trained models. Recent work has extended this paradigm to speaker identity unlear

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Convex-Neural RRT*: Fast and Reliable Learning-Guided Sampling for High-Quality Robot Path Planning

DGX agent

arXiv:2605.25006v1 Announce Type: cross Abstract: Sampling-based algorithms for robot path planning offer probabilistic completeness and strong empirical convergence properties across environments wit

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Counterfactual Explanations for Hypergraph Neural Networks

DGX agent

arXiv:2602.04360v2 Announce Type: replace-cross Abstract: Hypergraph neural networks (HGNNs) effectively model higher-order interactions in many real-world systems but remain difficult to interpret, l

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Courtroom Analogy: New Perspective on Uncertainty-Aware Classification

DGX agent

arXiv:2605.25616v1 Announce Type: new Abstract: Single-pass uncertainty quantification (UQ) methods for classification represent uncertainty by predicting a tractable distribution over the class proba

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Critical Organization of Deep Neural Networks, and p-Adic Statistical Field Theories

DGX agent

arXiv:2601.19070v2 Announce Type: replace Abstract: We rigorously study the thermodynamic limit of deep neural networks (DNNS) and recurrent neural networks (RNNs), assuming that the activation functi

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Cross-Domain Energy-Guided Diffusion Generation for Off-Dynamics Reinforcement Learning

DGX agent

arXiv:2605.24810v1 Announce Type: cross Abstract: Off-dynamics offline reinforcement learning seeks to learn a target-domain policy from a large source dataset and a limited target dataset under misma

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Cross-Domain Generalization Limits of Vision Foundation Models in Facial Deepfake Detection

DGX agent

arXiv:2605.24965v1 Announce Type: cross Abstract: The rapid evolution of generative models has enabled the creation of hyper-realistic facial deepfakes, exposing a critical vulnerability in modern dig

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CSP-Atlas: Concept-Specific Neural Circuits in a Sparse Python Transformer

DGX agent

arXiv:2605.24603v1 Announce Type: new Abstract: A sparse 8-layer code transformer develops dedicated neural circuitry for every Python construct tested, and that circuitry is organised by a clean comp

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents

DGX agent

arXiv:2605.25624v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven breakthroughs in domains such as math, tool-use, and software engineering, yet its exte

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CurveRL: Principled Distribution-Aware Context Reweighting for LLM Reasoning

DGX agent

arXiv:2605.24331v1 Announce Type: new Abstract: Context or prompt-level reweighting has emerged as a central algorithmic lever in Reinforcement Learning with Verified Rewards (RLVR) for improving the

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

CyberMaskQA: A Privacy-Aware Benchmark for Evaluating Large Language Models in Cybersecurity Question Answering

DGX agent

arXiv:2605.24765v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly applied to cybersecurity question answering (QA) for critical tasks such as incident response and vulner

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

D^2-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routing

DGX agent

arXiv:2605.25893v1 Announce Type: new Abstract: Despite the emergence of diffusion large language models (D-LLMs) as an alternative to autoregressive large language models (AR-LLMs), safety monitoring

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs

DGX agent

arXiv:2605.25188v1 Announce Type: new Abstract: Multi-agent LLM systems improve reasoning by combining outputs from multiple agents, but interaction-heavy methods can introduce error propagation and h

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Data-Specific Hyper-Parameter Design: A Paradigm Shift in Reservoir Computing

DGX agent

arXiv:2605.25221v1 Announce Type: cross Abstract: Reservoir computing typically relies on large, randomly generated reservoirs, enabling simple, often linear readouts. Over the past two decades, most

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Decision-Making with Lightweight Confidence-Aware Language Model for Autonomous Driving

DGX agent

arXiv:2605.25393v1 Announce Type: new Abstract: Large Language Models (LLMs) and Multimodal LLMs (MLLMs) have demonstrated immense potential in autonomous driving (AD) by offering human-like reasoning

model-releasesarxiv-cs-ro
26 May 2026
Model Releases

Decompose-and-Refine: Structured Legal Question Answering with Parametric Retrieval

DGX agent

arXiv:2605.24454v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong performance in the legal domain, demonstrating notable potential in Legal Question Answering (LQA). Howev

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Deployment-complete benchmarking

DGX agent

arXiv:2605.25997v1 Announce Type: new Abstract: Benchmarks increasingly guide deployment, procurement and scientific screening, yet a score supports only the response it records, not necessarily the d

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Directional Alignment Mitigates Reward Hacking in Reinforcement Learning for Language Models

DGX agent

arXiv:2605.25189v1 Announce Type: cross Abstract: Reward hacking arises when a model improves a proxy reward by exploiting shortcuts rather than solving the intended task. We study this failure mode t

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

DiscoverPhysics: Benchmarking LLMs for Out-of-the-Box Scientific Thinking

DGX agent

arXiv:2605.26087v1 Announce Type: cross Abstract: Frontier LLMs now perform strongly across a wide range of physics evaluations, but it is hard to disentangle genuine reasoning from recall of establis

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Distilling Game Code World Model Generation into Lightweight Large Language Models

DGX agent

arXiv:2605.24375v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown great ability in generating executable code from natural language, opening the possibility of automatically cons

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…263264265266267…472
Next →