AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,813 results
Model Releases

Bridging the Sim-to-Real Gap in Semiconductor Visual Program Synthesis via Input Binarization

DGX agent

arXiv:2606.02434v1 Announce Type: new Abstract: Precise parametric control over circuit geometry is essential for semiconductor inspection, yet obtaining sufficient real training data remains costly.

model-releasesarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Business Utility of Large Language Models as Exploratory Data Analysis Agents

DGX agent

arXiv:2606.00051v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in analytical workflows, but their suitability as exploratory data analysis (EDA) agents in busines

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CAFOSat: A Strongly Annotated Dataset for Infrastructure-Aware CAFO Mapping Using High-Resolution Imagery

DGX agent

arXiv:2606.00548v1 Announce Type: cross Abstract: Concentrated Animal Feeding Operations (CAFOs) play an important role in agricultural production but are also associated with environmental, public he

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures

DGX agent

arXiv:2505.24069v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are deployed on increasingly complex tasks that require multi-step decision-making. Understanding their algorithm

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM judges that check an agent’s work against rubric criter…

DGX agent

Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM judges that check an agent’s work against rubric criteria: they're used both in agent benchmarking and as reward si

model-releasesharrison-chase--x
2 Jun 2026
Model Releases

CART: Context-Anchored Recurrent Transformer -- A Parameter-Efficient Architecture with Learned Stability

DGX agent

arXiv:2606.01495v1 Announce Type: cross Abstract: We present CART (Context-Anchored Recurrent Transformer), a parameter-efficient language model that reuses a single shared core block R times across d

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CARTE: A Benchmark for Mapping Language Model Knowledge Across France

DGX agent

arXiv:2606.01995v1 Announce Type: new Abstract: We introduce CARTE 1 (Culturally Anchored Regional-Territorial Evaluation), a multiplechoice benchmark for evaluating the ability of large language mode

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CASTLE2026 Team WDL Technical Report

DGX agent

arXiv:2606.00712v1 Announce Type: new Abstract: The CASTLE Challenge @ EgoVis 2026 evaluates long-form egocentric video question answering over 600+ hours of multi-perspective recordings. Each four-ch

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Catch-Only-One: Non-Transferable Examples for Model-Specific Authorization

DGX agent

arXiv:2510.10982v2 Announce Type: replace-cross Abstract: Recent AI regulations increasingly emphasize the need for mechanisms that preserve the utility of data for AI innovation while preventing misu

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Chameleon: Style-Content Disentangled Framework for Cross-Domain Object Compositing

DGX agent

arXiv:2606.01079v1 Announce Type: new Abstract: Image compositing aims to seamlessly insert a foreground object into a background image, and recent advances in diffusion models have significantly enha

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Characterization of Multi-Model Agentic AI Systems on General Tasks via Trace-Driven Simulation

DGX agent

arXiv:2606.01725v1 Announce Type: new Abstract: Agentic AI completes tasks through iterative planning, tool use, and reasoning based on observed outcomes. Despite its popularity, its system-level beha

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ChartArena: Benchmarking Chart Parsing across Languages, Scenarios, and Formats

DGX agent

arXiv:2606.01348v1 Announce Type: new Abstract: Charts are a primary medium for conveying quantitative and relational information, yet systematically evaluating chart parsing models remains difficult.

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Child-directed speech facilitates production, not comprehension, in BabyLMs

DGX agent

arXiv:2606.01045v1 Announce Type: new Abstract: Recent studies suggest that child-directed speech is not conducive to language learning in BabyLMs. However, current evaluations focus predominantly on

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs

DGX agent

arXiv:2606.00898v1 Announce Type: new Abstract: Large language models systematically hallucinate legal citations -- fabricating statute references, citing repealed provisions, and confusing jurisdicti

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CityTrajBench: A Unified Benchmark for City-Scale Vehicle Trajectory Generation

DGX agent

arXiv:2606.02287v1 Announce Type: cross Abstract: Urban trajectory generation is a fundamental task for transportation simulation, urban planning, and mobility analytics. However, systematic compariso

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Claude Agent now runs inside Devin Desktop. Start Claude Agent sessions in Devin Desktop, coordinate them with your other agents, and keep c…

DGX agent

Claude Agent now runs inside Devin Desktop. Start Claude Agent sessions in Devin Desktop, coordinate them with your other agents, and keep context shared from a single command center. Learn more about

model-releaseswindsurf--x
2 Jun 2026
Model Releases

Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs

DGX agent

arXiv:2603.24511v2 Announce Type: replace-cross Abstract: We show that AI agents are capable of discovering novel algorithms for adversarial attacks against LLMs, advancing the state of the art on whi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

DGX agent

arXiv:2606.01494v1 Announce Type: cross Abstract: Agent skills extend AI agents with reusable instructions, tools, scripts, references, and workflows, establishing a security boundary distinct from bo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents

DGX agent

arXiv:2606.02568v1 Announce Type: new Abstract: Clinical practice is not the selection of an answer from enumerated options: a physician gathers heterogeneous information incrementally and commits to

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection

DGX agent

arXiv:2606.00101v1 Announce Type: cross Abstract: With the rapid advancement of artificial intelligence generated content (AIGC) technologies, video forgery has become increasingly prevalent, posing n

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CodeCytos: AI-assisted spatial molecular imaging analysis via code-augmented agent action space

DGX agent

arXiv:2606.00472v1 Announce Type: cross Abstract: Conventional tissue image analysis software provides foundational capabilities for cellular analysis, including segmentation, basic morphological feat

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Codex for every role, tool, and workflow

DGX agent

OpenAI's Codex is an AI system that translates natural language instructions into code, designed to assist users across different roles, tools, and workflows. It enables developers, non-technical user

model-releasesopenai
2 Jun 2026
Model Releases

Codex is becoming a productivity tool for everyone

DGX agent

OpenAI's Codex is evolving beyond code generation to become a general productivity tool accessible to non-programmers for knowledge work tasks. The tool leverages large language models to assist with

model-releasesopenai
2 Jun 2026
Model Releases

Collaborative and Efficient Fine-tuning: Leveraging Task Similarity

DGX agent

arXiv:2602.07218v2 Announce Type: replace-cross Abstract: Adaptability has been regarded as a central feature in the foundation models, enabling them to effectively acclimate to unseen downstream task

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models

DGX agent

arXiv:2606.00658v1 Announce Type: cross Abstract: Large video diffusion models achieve strong visual quality but remain expensive to deploy because each sample requires many denoising steps and a larg

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CoMIC: Collaborative Memory and Insights Circulation for Long-Horizon LLM Agents in Cloud-Edge Systems

DGX agent

arXiv:2606.00756v1 Announce Type: new Abstract: Deploying lightweight Large Language Model (LLM) agents on edge servers can reduce latency and move agentic services closer to users, but resource-const

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Community-Aware Assessment of Social Textual Engagement and Resonance: A Human-Centric Perspective on User-Generated Content Evaluation

DGX agent

arXiv:2606.01897v1 Announce Type: new Abstract: Traditional Video Quality Assessment (VQA) focuses narrowly on aesthetic fidelity, overlooking the complex social dynamics that define quality in User-G

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Completion at the Boundary (CaB): Deployable Switching with Completion-Aware Control under Limited Calibration

DGX agent

arXiv:2606.00145v1 Announce Type: cross Abstract: Vision-language-action (VLA) agents can execute natural-language instructions, yet deployed systems still lack an operational interface: deciding when

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Connecting AI agents with unstructured data using Google Cloud Storage MCP Servers

DGX agent

Google Cloud Storage (GCS) is a foundational component of the modern agentic tech stack and the preferred home for unstructured data at scale. As enterprises deploy agents in production, the critical

model-releasesgoogle-cloud-ai
2 Jun 2026
Model Releases

Connecting the Dots: Benchmarking Reflective Memory in Long-Horizon Dialogue

DGX agent

arXiv:2606.01223v1 Announce Type: cross Abstract: Despite substantial progress in long-context modeling, existing benchmarks remain confined to factual memory for explicit recall, failing to measure t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Consistency evaluation of benchmarks used for causal discovery

DGX agent

arXiv:2606.01789v1 Announce Type: new Abstract: In graphical causal model, causal discovery aims to construct a causal graph based on numerical data and domain knowledge in plain text. However, the ev

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Consistent and Distinctive: LLM Benchmark Efficiency via Maximum Independent Set Prompt Selection on Similarity Graphs

DGX agent

arXiv:2606.01400v1 Announce Type: cross Abstract: Evaluating large language models (LLMs) across comprehensive benchmarks is expensive and time-consuming. We propose a graph-based prompt selection fra

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem

DGX agent

arXiv:2603.16572v2 Announce Type: replace-cross Abstract: Agent skills extend local AI agents, such as Claude Code and OpenClaw, with additional functionality. Their growing popularity has led to dedi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Continual Learning involves engineering whole systems including Efficient Verifiers to make RL/fine-tuning and running evaluations much chea…

DGX agent

Continual Learning involves engineering whole systems including Efficient Verifiers to make RL/fine-tuning and running evaluations much cheaper at scale! some initial work we’re releasing from LangCha

model-releasesharrison-chase--x
2 Jun 2026
Model Releases

ContinuousBench: Can Differentially Private Synthetic Text Improve Capabilities?

DGX agent

arXiv:2606.01849v1 Announce Type: cross Abstract: Differentially private (DP) text synthesis promises to unlock sensitive corpora for model training, but it remains unclear whether DP synthetic data t

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Controllable Value Alignment in Large Language Models through Neuron-Level Editing

DGX agent

arXiv:2602.07356v2 Announce Type: replace Abstract: Aligning large language models (LLMs) with human values has become increasingly important as their influence on human behavior and decision-making e

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Correcting Gradient-Based Circuit Localization via Interaction-Aware Backpropagation

DGX agent

arXiv:2505.17630v4 Announce Type: replace Abstract: Circuit localization methods aim to identify the subset of model components responsible for specific behaviors in large language models, enabling de

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRAB-Bench: Evaluating LLM Agents under Complex Task Dependencies and Human-aligned User Simulation

DGX agent

arXiv:2606.01815v1 Announce Type: new Abstract: Evaluating LLM agents in realistic service scenarios requires complex task dependencies, imperfect user behavior, and an evaluation that accommodates mu

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRAM: Centroid-Routing and Adaptive MoE for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2606.02502v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) unify heterogeneous vision-language tasks under a shared generative framework via instruction tuning, yet real-

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CRMA: A Spectrally-Bounded Backbone for Modular Continual Fine-Tuning of LLMs

DGX agent

arXiv:2606.00382v1 Announce Type: new Abstract: Sequential fine-tuning of large language models forces a choice: let the shared substrate keep learning and accept catastrophic forgetting, or freeze it

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Cross-Environment Neural Reranking for Sample-Efficient Action Selection in Text-Based Agents

DGX agent

arXiv:2606.02204v1 Announce Type: new Abstract: Large language model agents achieve strong performance on text-based benchmarks but incur prohibitive inference costs, motivating the use of compact neu

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Cross-Generational Transfer of Adversarial Attacks Reveals Non-Monotonic Safety Alignment in LLMs

DGX agent

arXiv:2606.00813v1 Announce Type: cross Abstract: Safety alignment in LLMs does not improve monotonically across model generations. Studying four generations of Google's Gemma family (7B-31B) with qua

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CSRP: Chain-of-Thought Reasoning for Chinese Text Correction via Reinforcement Learning with Efficiency-Aware Rewards

DGX agent

arXiv:2606.00020v1 Announce Type: cross Abstract: Large Language Model (LLM) based Chinese Grammatical Error Correction (CGEC) systems face two critical challenges: general-purpose models lack special

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs

DGX agent

arXiv:2606.01879v1 Announce Type: new Abstract: Existing research largely reduces cultural intelligence in LLMs to a knowledge-level problem, overlooking whether models can effectively utilize their a

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CV-Arena: An Open Benchmark for Instructional Computer Vision Problem Solving with Human-AI Collaborative Preferences

DGX agent

arXiv:2606.00931v1 Announce Type: cross Abstract: Instruction-guided image editing is becoming a general interface for visual work, yet existing benchmarks still focus largely on narrow appearance edi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAG-MoE: From Simple Mixture to Structural Aggregation in Mixture-of-Experts

DGX agent

arXiv:2606.01062v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become a leading approach for decoupling parameter count from computational cost in large language models, yet effe

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAG-Plan: Generating Directed Acyclic Dependency Graphs for Dual-Arm Cooperative Planning

DGX agent

arXiv:2406.09953v4 Announce Type: replace-cross Abstract: Dual-arm robots promise greater efficiency but require planning for complex tasks with nonlinear sub-task dependencies. Current methods using

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DASH: Dual-Branch Score Distillation for Guidance-Calibrated Compact Diffusion Models

DGX agent

arXiv:2606.00798v1 Announce Type: cross Abstract: Parameter compression of class-conditional diffusion models reveals an underexplored limitation in output-level distillation: the unconditional score

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…228229230231232…476
Next →