AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Benchmarking Multimodal LLMs on Code Generation for Complex Interactive Webpages

DGX agent

arXiv:2606.00154v1 Announce Type: cross Abstract: Recent advancements in multimodal large language models (MLLMs) have achieved remarkable progress in multimodal reasoning and code generation, catalyz

model-releasesarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Benchmarking Recursive-Collapse Warning Claims Under Matched False-Positive Control

DGX agent

arXiv:2606.00329v1 Announce Type: cross Abstract: Recursive systems can enter collapse-like regimes -- self-reinforcing amplification, persistent recursion, and narrowing diversity that mask accelerat

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Benchmarking Security Risk Detection and Verification in Open Agentic Skill Ecosystems

DGX agent

arXiv:2606.00925v1 Announce Type: cross Abstract: Open agent platforms allow community contributors to publish reusable skills that agents can invoke at runtime. This extensibility also creates a supp

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Benchmarking Waitlist Mortality Prediction in Heart Transplantation Through Time-to-Event Modeling using New Longitudinal UNOS Dataset

DGX agent

arXiv:2507.07339v2 Announce Type: replace-cross Abstract: Decisions about managing patients on the heart transplant waitlist are currently made by committees of doctors who consider multiple factors,

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Benchmarks for Vision-Language Models in Urban Perception Should Be Reliability-Aware and Negotiated

DGX agent

arXiv:2606.00871v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used to generate structured descriptions of street-level imagery for tasks such as streetscape auditing

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Better with Experience: Self-Evolving LLM Agents for Evidence-Grounded Health Community Notes

DGX agent

arXiv:2606.02215v1 Announce Type: new Abstract: Large Language Model (LLM)-augmented Community Notes offer a scalable path for timely, evidence-grounded correction of health misinformation on social p

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Beware of the Batch Size: Hyperparameter Bias in Evaluating LoRA

DGX agent

arXiv:2602.09492v2 Announce Type: replace-cross Abstract: Low-rank adaptation (LoRA) is a standard approach for fine-tuning large language models, yet its many variants report conflicting empirical ga

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Beyond ell_2-norm and ell_infty-norm: A Curvature-Inspired ell_p-Norm Scheme for Deep Neural Networks

DGX agent

arXiv:2606.02078v1 Announce Type: new Abstract: The existing optimizers for deep neural networks (DNNs) typically rely on either the ell_2 norm or the ell_infty norm, resulting in optimizers that do n

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization

DGX agent

arXiv:2606.02300v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across diverse domains, yet personalizing their outputs to individual users remai

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Beyond Rigid: Benchmarking Non-Rigid Video Editing

DGX agent

arXiv:2601.18340v2 Announce Type: replace Abstract: As video generation models are increasingly expected to manipulate physical dynamics, there is a growing need to move evaluation beyond appearance f

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

DGX agent

arXiv:2603.19453v2 Announce Type: replace Abstract: We study LLM policy synthesis: using a language model to iteratively generate programmatic agent policies for multi-agent environments. Rather than

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Beyond Semantic Understanding: Preserving Collaborative Frequency Components in LLM-based Recommendation

DGX agent

arXiv:2508.10312v2 Announce Type: replace Abstract: Recommender systems in concert with Large Language Models (LLMs) present promising avenues for generating semantically-informed recommendations. How

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Beyond Static Gaussians: An Empirical Investigation of Architectural Paradigms for Dynamic 3D Scene Reconstruction

DGX agent

arXiv:2606.00452v1 Announce Type: new Abstract: Dynamic scene reconstruction via 3D Gaussian Splatting (3DGS) has emerged as a compelling approach for representing evolving environments, yet understan

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Beyond Text and Tables: Vision-Language Model Integration in ComProScanner for Extracting Materials Data from Scientific Figures with High Accuracy

DGX agent

arXiv:2606.00065v1 Announce Type: cross Abstract: Automated extraction of materials composition-property data from scientific literature has advanced considerably with the development of large languag

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Beyond the Simplex: Balanced Prototype Geometry for Scorer-Agnostic Open-Set Recognition

DGX agent

arXiv:2606.01883v1 Announce Type: cross Abstract: Open-set recognition (OSR) requires a classifier to reject inputs from unseen classes which is essential in safety-critical settings such as medical i

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

BLISS: A Lightweight Bilevel Influence Scoring Method for Data Selection in Language Model Pretraining

DGX agent

arXiv:2510.06048v4 Announce Type: replace Abstract: Effective data selection is essential for pretraining large language models (LLMs), enhancing efficiency and improving generalization to downstream

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention

DGX agent

arXiv:2512.10414v2 Announce Type: replace Abstract: Recently, reinforcement learning (RL) has become a common choice in enhancing the reasoning capabilities of vision-language models (VLMs). Consideri

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Boundary-Protection W8A8 HiFloat8 Quantization for Large-Scale Text-to-Video Diffusion Transformers

DGX agent

arXiv:2606.00957v1 Announce Type: new Abstract: We present a post-training quantization (PTQ) approach for Wan2.1-T2V-14B, a 14-billion-parameter text-to-video diffusion transformer, targeting the W8A

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

BraveGuard: From Open-World Threats to Safer Computer-Use Agents

DGX agent

arXiv:2606.01166v1 Announce Type: cross Abstract: Computer-use agents extend language models from text generation to sustained interaction with files, terminals, browsers, and external tools. This shi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory

DGX agent

arXiv:2606.01385v1 Announce Type: cross Abstract: Software architecture design is a critical yet inherently complex and knowledge-intensive phase that requires balancing competing quality attributes a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Bridging the Sim-to-Real Gap in Semiconductor Visual Program Synthesis via Input Binarization

DGX agent

arXiv:2606.02434v1 Announce Type: new Abstract: Precise parametric control over circuit geometry is essential for semiconductor inspection, yet obtaining sufficient real training data remains costly.

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Business Utility of Large Language Models as Exploratory Data Analysis Agents

DGX agent

arXiv:2606.00051v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in analytical workflows, but their suitability as exploratory data analysis (EDA) agents in busines

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CAFOSat: A Strongly Annotated Dataset for Infrastructure-Aware CAFO Mapping Using High-Resolution Imagery

DGX agent

arXiv:2606.00548v1 Announce Type: cross Abstract: Concentrated Animal Feeding Operations (CAFOs) play an important role in agricultural production but are also associated with environmental, public he

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Can LLMs Reason Structurally? Benchmarking via the Lens of Data Structures

DGX agent

arXiv:2505.24069v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are deployed on increasingly complex tasks that require multi-step decision-making. Understanding their algorithm

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CART: Context-Anchored Recurrent Transformer -- A Parameter-Efficient Architecture with Learned Stability

DGX agent

arXiv:2606.01495v1 Announce Type: cross Abstract: We present CART (Context-Anchored Recurrent Transformer), a parameter-efficient language model that reuses a single shared core block R times across d

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CARTE: A Benchmark for Mapping Language Model Knowledge Across France

DGX agent

arXiv:2606.01995v1 Announce Type: new Abstract: We introduce CARTE 1 (Culturally Anchored Regional-Territorial Evaluation), a multiplechoice benchmark for evaluating the ability of large language mode

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CASTLE2026 Team WDL Technical Report

DGX agent

arXiv:2606.00712v1 Announce Type: new Abstract: The CASTLE Challenge @ EgoVis 2026 evaluates long-form egocentric video question answering over 600+ hours of multi-perspective recordings. Each four-ch

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Catch-Only-One: Non-Transferable Examples for Model-Specific Authorization

DGX agent

arXiv:2510.10982v2 Announce Type: replace-cross Abstract: Recent AI regulations increasingly emphasize the need for mechanisms that preserve the utility of data for AI innovation while preventing misu

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Chameleon: Style-Content Disentangled Framework for Cross-Domain Object Compositing

DGX agent

arXiv:2606.01079v1 Announce Type: new Abstract: Image compositing aims to seamlessly insert a foreground object into a background image, and recent advances in diffusion models have significantly enha

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Characterization of Multi-Model Agentic AI Systems on General Tasks via Trace-Driven Simulation

DGX agent

arXiv:2606.01725v1 Announce Type: new Abstract: Agentic AI completes tasks through iterative planning, tool use, and reasoning based on observed outcomes. Despite its popularity, its system-level beha

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ChartArena: Benchmarking Chart Parsing across Languages, Scenarios, and Formats

DGX agent

arXiv:2606.01348v1 Announce Type: new Abstract: Charts are a primary medium for conveying quantitative and relational information, yet systematically evaluating chart parsing models remains difficult.

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Child-directed speech facilitates production, not comprehension, in BabyLMs

DGX agent

arXiv:2606.01045v1 Announce Type: new Abstract: Recent studies suggest that child-directed speech is not conducive to language learning in BabyLMs. However, current evaluations focus predominantly on

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs

DGX agent

arXiv:2606.00898v1 Announce Type: new Abstract: Large language models systematically hallucinate legal citations -- fabricating statute references, citing repealed provisions, and confusing jurisdicti

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

CityTrajBench: A Unified Benchmark for City-Scale Vehicle Trajectory Generation

DGX agent

arXiv:2606.02287v1 Announce Type: cross Abstract: Urban trajectory generation is a fundamental task for transportation simulation, urban planning, and mobility analytics. However, systematic compariso

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs

DGX agent

arXiv:2603.24511v2 Announce Type: replace-cross Abstract: We show that AI agents are capable of discovering novel algorithms for adversarial attacks against LLMs, advancing the state of the art on whi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

DGX agent

arXiv:2606.01494v1 Announce Type: cross Abstract: Agent skills extend AI agents with reusable instructions, tools, scripts, references, and workflows, establishing a security boundary distinct from bo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents

DGX agent

arXiv:2606.02568v1 Announce Type: new Abstract: Clinical practice is not the selection of an answer from enumerated options: a physician gathers heterogeneous information incrementally and commits to

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection

DGX agent

arXiv:2606.00101v1 Announce Type: cross Abstract: With the rapid advancement of artificial intelligence generated content (AIGC) technologies, video forgery has become increasingly prevalent, posing n

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CodeCytos: AI-assisted spatial molecular imaging analysis via code-augmented agent action space

DGX agent

arXiv:2606.00472v1 Announce Type: cross Abstract: Conventional tissue image analysis software provides foundational capabilities for cellular analysis, including segmentation, basic morphological feat

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Collaborative and Efficient Fine-tuning: Leveraging Task Similarity

DGX agent

arXiv:2602.07218v2 Announce Type: replace-cross Abstract: Adaptability has been regarded as a central feature in the foundation models, enabling them to effectively acclimate to unseen downstream task

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models

DGX agent

arXiv:2606.00658v1 Announce Type: cross Abstract: Large video diffusion models achieve strong visual quality but remain expensive to deploy because each sample requires many denoising steps and a larg

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CoMIC: Collaborative Memory and Insights Circulation for Long-Horizon LLM Agents in Cloud-Edge Systems

DGX agent

arXiv:2606.00756v1 Announce Type: new Abstract: Deploying lightweight Large Language Model (LLM) agents on edge servers can reduce latency and move agentic services closer to users, but resource-const

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Community-Aware Assessment of Social Textual Engagement and Resonance: A Human-Centric Perspective on User-Generated Content Evaluation

DGX agent

arXiv:2606.01897v1 Announce Type: new Abstract: Traditional Video Quality Assessment (VQA) focuses narrowly on aesthetic fidelity, overlooking the complex social dynamics that define quality in User-G

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Completion at the Boundary (CaB): Deployable Switching with Completion-Aware Control under Limited Calibration

DGX agent

arXiv:2606.00145v1 Announce Type: cross Abstract: Vision-language-action (VLA) agents can execute natural-language instructions, yet deployed systems still lack an operational interface: deciding when

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Connecting the Dots: Benchmarking Reflective Memory in Long-Horizon Dialogue

DGX agent

arXiv:2606.01223v1 Announce Type: cross Abstract: Despite substantial progress in long-context modeling, existing benchmarks remain confined to factual memory for explicit recall, failing to measure t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Consistency evaluation of benchmarks used for causal discovery

DGX agent

arXiv:2606.01789v1 Announce Type: new Abstract: In graphical causal model, causal discovery aims to construct a causal graph based on numerical data and domain knowledge in plain text. However, the ev

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Consistent and Distinctive: LLM Benchmark Efficiency via Maximum Independent Set Prompt Selection on Similarity Graphs

DGX agent

arXiv:2606.01400v1 Announce Type: cross Abstract: Evaluating large language models (LLMs) across comprehensive benchmarks is expensive and time-consuming. We propose a graph-based prompt selection fra

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem

DGX agent

arXiv:2603.16572v2 Announce Type: replace-cross Abstract: Agent skills extend local AI agents, such as Claude Code and OpenClaw, with additional functionality. Their growing popularity has led to dedi

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…164165166167168…361
Next →