AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,797 results
Model Releases

Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs

DGX agent

arXiv:2603.24511v2 Announce Type: replace-cross Abstract: We show that AI agents are capable of discovering novel algorithms for adversarial attacks against LLMs, advancing the state of the art on whi

model-releasesarxiv-cs-ai
2 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

DGX agent

arXiv:2606.01494v1 Announce Type: cross Abstract: Agent skills extend AI agents with reusable instructions, tools, scripts, references, and workflows, establishing a security boundary distinct from bo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents

DGX agent

arXiv:2606.02568v1 Announce Type: new Abstract: Clinical practice is not the selection of an answer from enumerated options: a physician gathers heterogeneous information incrementally and commits to

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection

DGX agent

arXiv:2606.00101v1 Announce Type: cross Abstract: With the rapid advancement of artificial intelligence generated content (AIGC) technologies, video forgery has become increasingly prevalent, posing n

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CodeCytos: AI-assisted spatial molecular imaging analysis via code-augmented agent action space

DGX agent

arXiv:2606.00472v1 Announce Type: cross Abstract: Conventional tissue image analysis software provides foundational capabilities for cellular analysis, including segmentation, basic morphological feat

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Codex for every role, tool, and workflow

DGX agent

OpenAI's Codex is an AI system that translates natural language instructions into code, designed to assist users across different roles, tools, and workflows. It enables developers, non-technical user

model-releasesopenai
2 Jun 2026
Model Releases

Codex is becoming a productivity tool for everyone

DGX agent

OpenAI's Codex is evolving beyond code generation to become a general productivity tool accessible to non-programmers for knowledge work tasks. The tool leverages large language models to assist with

model-releasesopenai
2 Jun 2026
Model Releases

Collaborative and Efficient Fine-tuning: Leveraging Task Similarity

DGX agent

arXiv:2602.07218v2 Announce Type: replace-cross Abstract: Adaptability has been regarded as a central feature in the foundation models, enabling them to effectively acclimate to unseen downstream task

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models

DGX agent

arXiv:2606.00658v1 Announce Type: cross Abstract: Large video diffusion models achieve strong visual quality but remain expensive to deploy because each sample requires many denoising steps and a larg

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CoMIC: Collaborative Memory and Insights Circulation for Long-Horizon LLM Agents in Cloud-Edge Systems

DGX agent

arXiv:2606.00756v1 Announce Type: new Abstract: Deploying lightweight Large Language Model (LLM) agents on edge servers can reduce latency and move agentic services closer to users, but resource-const

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Community-Aware Assessment of Social Textual Engagement and Resonance: A Human-Centric Perspective on User-Generated Content Evaluation

DGX agent

arXiv:2606.01897v1 Announce Type: new Abstract: Traditional Video Quality Assessment (VQA) focuses narrowly on aesthetic fidelity, overlooking the complex social dynamics that define quality in User-G

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Completion at the Boundary (CaB): Deployable Switching with Completion-Aware Control under Limited Calibration

DGX agent

arXiv:2606.00145v1 Announce Type: cross Abstract: Vision-language-action (VLA) agents can execute natural-language instructions, yet deployed systems still lack an operational interface: deciding when

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Connecting AI agents with unstructured data using Google Cloud Storage MCP Servers

DGX agent

Google Cloud Storage (GCS) is a foundational component of the modern agentic tech stack and the preferred home for unstructured data at scale. As enterprises deploy agents in production, the critical

model-releasesgoogle-cloud-ai
2 Jun 2026
Model Releases

Connecting the Dots: Benchmarking Reflective Memory in Long-Horizon Dialogue

DGX agent

arXiv:2606.01223v1 Announce Type: cross Abstract: Despite substantial progress in long-context modeling, existing benchmarks remain confined to factual memory for explicit recall, failing to measure t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Consistency evaluation of benchmarks used for causal discovery

DGX agent

arXiv:2606.01789v1 Announce Type: new Abstract: In graphical causal model, causal discovery aims to construct a causal graph based on numerical data and domain knowledge in plain text. However, the ev

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Consistent and Distinctive: LLM Benchmark Efficiency via Maximum Independent Set Prompt Selection on Similarity Graphs

DGX agent

arXiv:2606.01400v1 Announce Type: cross Abstract: Evaluating large language models (LLMs) across comprehensive benchmarks is expensive and time-consuming. We propose a graph-based prompt selection fra

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem

DGX agent

arXiv:2603.16572v2 Announce Type: replace-cross Abstract: Agent skills extend local AI agents, such as Claude Code and OpenClaw, with additional functionality. Their growing popularity has led to dedi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Continual Learning involves engineering whole systems including Efficient Verifiers to make RL/fine-tuning and running evaluations much chea…

DGX agent

Continual Learning involves engineering whole systems including Efficient Verifiers to make RL/fine-tuning and running evaluations much cheaper at scale! some initial work we’re releasing from LangCha

model-releasesharrison-chase--x
2 Jun 2026
Model Releases

ContinuousBench: Can Differentially Private Synthetic Text Improve Capabilities?

DGX agent

arXiv:2606.01849v1 Announce Type: cross Abstract: Differentially private (DP) text synthesis promises to unlock sensitive corpora for model training, but it remains unclear whether DP synthetic data t

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Controllable Value Alignment in Large Language Models through Neuron-Level Editing

DGX agent

arXiv:2602.07356v2 Announce Type: replace Abstract: Aligning large language models (LLMs) with human values has become increasingly important as their influence on human behavior and decision-making e

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Correcting Gradient-Based Circuit Localization via Interaction-Aware Backpropagation

DGX agent

arXiv:2505.17630v4 Announce Type: replace Abstract: Circuit localization methods aim to identify the subset of model components responsible for specific behaviors in large language models, enabling de

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRAB-Bench: Evaluating LLM Agents under Complex Task Dependencies and Human-aligned User Simulation

DGX agent

arXiv:2606.01815v1 Announce Type: new Abstract: Evaluating LLM agents in realistic service scenarios requires complex task dependencies, imperfect user behavior, and an evaluation that accommodates mu

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRAM: Centroid-Routing and Adaptive MoE for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2606.02502v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) unify heterogeneous vision-language tasks under a shared generative framework via instruction tuning, yet real-

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRMA: A Spectrally-Bounded Backbone for Modular Continual Fine-Tuning of LLMs

DGX agent

arXiv:2606.00382v1 Announce Type: new Abstract: Sequential fine-tuning of large language models forces a choice: let the shared substrate keep learning and accept catastrophic forgetting, or freeze it

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Cross-Environment Neural Reranking for Sample-Efficient Action Selection in Text-Based Agents

DGX agent

arXiv:2606.02204v1 Announce Type: new Abstract: Large language model agents achieve strong performance on text-based benchmarks but incur prohibitive inference costs, motivating the use of compact neu

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Cross-Generational Transfer of Adversarial Attacks Reveals Non-Monotonic Safety Alignment in LLMs

DGX agent

arXiv:2606.00813v1 Announce Type: cross Abstract: Safety alignment in LLMs does not improve monotonically across model generations. Studying four generations of Google's Gemma family (7B-31B) with qua

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CSRP: Chain-of-Thought Reasoning for Chinese Text Correction via Reinforcement Learning with Efficiency-Aware Rewards

DGX agent

arXiv:2606.00020v1 Announce Type: cross Abstract: Large Language Model (LLM) based Chinese Grammatical Error Correction (CGEC) systems face two critical challenges: general-purpose models lack special

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs

DGX agent

arXiv:2606.01879v1 Announce Type: new Abstract: Existing research largely reduces cultural intelligence in LLMs to a knowledge-level problem, overlooking whether models can effectively utilize their a

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CV-Arena: An Open Benchmark for Instructional Computer Vision Problem Solving with Human-AI Collaborative Preferences

DGX agent

arXiv:2606.00931v1 Announce Type: cross Abstract: Instruction-guided image editing is becoming a general interface for visual work, yet existing benchmarks still focus largely on narrow appearance edi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAG-MoE: From Simple Mixture to Structural Aggregation in Mixture-of-Experts

DGX agent

arXiv:2606.01062v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become a leading approach for decoupling parameter count from computational cost in large language models, yet effe

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAG-Plan: Generating Directed Acyclic Dependency Graphs for Dual-Arm Cooperative Planning

DGX agent

arXiv:2406.09953v4 Announce Type: replace-cross Abstract: Dual-arm robots promise greater efficiency but require planning for complex tasks with nonlinear sub-task dependencies. Current methods using

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DASH: Dual-Branch Score Distillation for Guidance-Calibrated Compact Diffusion Models

DGX agent

arXiv:2606.00798v1 Announce Type: cross Abstract: Parameter compression of class-conditional diffusion models reveals an underexplored limitation in output-level distillation: the unconditional score

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAStatFormer: A Hybrid Multibranch Transformer with Statistical Feature Integration for DAS-Based Pattern Recognitions

DGX agent

arXiv:2606.00081v1 Announce Type: cross Abstract: Distributed Acoustic Sensing (DAS) enables large-scale monitoring through optical fibers, but its high dimensionality and complex spatio-temporal patt

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Data agents don't fail at writing SQL. They fail at knowing your business. Schemas show you the columns, but they don't tell you which view …

DGX agent

Data agents don't fail at writing SQL. They fail at knowing your business. Schemas show you the columns, but they don't tell you which view is canonical for ARR, how often each metric updates, or whic

model-releasespinecone--x
2 Jun 2026
Model Releases

Data Collection for Training Quality-Control AI in Carpet Manufacturing

DGX agent

arXiv:2606.01023v1 Announce Type: cross Abstract: Visual inspection remains the dominant quality-control practice in woven and tufted carpet production, yet it is slow, subjective, and inconsistent at

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

datasette-agent-micropython 0.1a0

DGX agent

Release: datasette-agent-micropython 0.1a0 I want Datasette Agent to be able to generate and execute Python code safely. This alpha is looking promising so far. GPT-5.5 has so far failed to break out

model-releasessimon-willison
2 Jun 2026
Model Releases

Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging

DGX agent

arXiv:2606.01717v1 Announce Type: new Abstract: Instruction tuning aligns large language models, including multimodal ones, with diverse user intents, but scaling to heterogeneous mixtures is hindered

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Decision-Focused On-Policy Learning for Contextual Linear Optimization with Partial Feedback

DGX agent

arXiv:2606.01081v1 Announce Type: new Abstract: Decision-focused learning (DFL) trains predictive models by optimizing downstream decision quality rather than standalone prediction accuracy. For conte

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

DECK: A Consistency x Confidence Taxonomy of LLM Hallucinations

DGX agent

arXiv:2606.02289v1 Announce Type: new Abstract: Existing hallucination taxonomies classify LLM errors by what is wrong with the output -- memorised misconceptions, reasoning failures, fluent fabricati

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Deep Research as Rubric for Reinforcement Learning

DGX agent

arXiv:2606.01091v1 Announce Type: new Abstract: Open-ended reasoning and long-form generation tasks lack reliable automatic verification signals for reward-based policy optimization. Rubrics offer a p

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Deformable Wiener Filter for Future Video Coding

DGX agent

arXiv:2606.01576v1 Announce Type: new Abstract: In-loop filters have attracted increasing attention due to the remarkable noise-reduction capability in the hybrid video coding framework. However, the

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs

DGX agent

arXiv:2606.01710v1 Announce Type: new Abstract: Vision-Language models (VLMs), such as CLIP, achieve powerful zero-shot classification. However, their predictions remain sensitive to spurious correlat

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Design-MLLM: A Reinforcement Alignment Framework for Verifiable and Aesthetic Interior Design

DGX agent

arXiv:2603.13312v2 Announce Type: replace-cross Abstract: Interior design is a requirements-to-visual-plan generation process that must simultaneously satisfy verifiable spatial feasibility and compar

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?

DGX agent

arXiv:2505.16915v3 Announce Type: replace-cross Abstract: While recent Text-to-Image (T2I) models show impressive capabilities in synthesizing images from brief descriptions, they struggle with the lo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Diagnosing LLM Arbitration Behavior over Pre-evidence Epistemic States in RAG-based Fact-Checking

DGX agent

arXiv:2606.01120v1 Announce Type: new Abstract: In RAG-based fact-checking, LLMs are increasingly used as verifiers to check given claims against retrieved evidence. Their parametric knowledge can ind

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Differentially Private Datastore Generation for Retrieval-Augmented Inference

DGX agent

arXiv:2606.01413v1 Announce Type: cross Abstract: It is crucial for modern on-device AI systems that rely on retrieval-augmented inference to release and share datastores without compromising individu

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

DINO-GFSA: Geo-Localization via Semantic Gated Fusion and Mamba-based Sequential Aggregation

DGX agent

arXiv:2606.00784v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) is critical for Unmanned Aerial Vehicle (UAV) self-positioning and target localization in GNSS-denied environments. H

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Disentanglement-Based Equivariant Learning for Compositional VQA

DGX agent

arXiv:2606.02168v1 Announce Type: new Abstract: Compositional visual question answering (VQA) represents a challenging yet fundamental task that requires models to comprehend novel combinations of pre

model-releasesarxiv-cs-cv
2 Jun 2026
← Previous
1…227228229230231…475
Next →