AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,555 results
Model Releases

Can Large Language Models Implement Agent-Based Models? An ODD-based Replication Study

DGX agent

arXiv:2602.10140v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can now synthesize non-trivial executable code from textual descriptions, raising an important question: can LLMs

model-releasesarxiv-cs-ai
1 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

CareGuardAI: Context-Aware Multi-Agent Guardrails for Clinical Safety & Hallucination Mitigation in Patient-Facing LLMs

DGX agent

arXiv:2604.26959v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into patient-facing healthcare systems offers significant potential to improve access to medical information.

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

CausalCompass: Evaluating the Robustness of Time-Series Causal Discovery in Misspecified Scenarios

DGX agent

arXiv:2602.07915v2 Announce Type: replace-cross Abstract: Causal discovery from time series is a fundamental task in machine learning. However, its widespread adoption is hindered by a reliance on unt

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Characterizing the Consistency of the Emergent Misalignment Persona

DGX agent

arXiv:2604.28082v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) on narrowly misaligned data generalizes to broadly misaligned behavior, a phenomenon termed emergent misalignme

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ChipLingo: A Systematic Training Framework for Large Language Models in EDA

DGX agent

arXiv:2604.27415v1 Announce Type: new Abstract: With the rapid advancement of semiconductor technology, Electronic Design Automation (EDA) has become an increasingly knowledge-intensive and document-d

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

CL-bench Life: Can Language Models Learn from Real-Life Context?

DGX agent

arXiv:2604.27043v1 Announce Type: new Abstract: Today's AI assistants such as OpenClaw are designed to handle context effectively, making context learning an increasingly important capability for mode

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows

DGX agent

arXiv:2604.28139v1 Announce Type: cross Abstract: LLM agents are expected to complete end-to-end units of work across software tools, business services, and local workspaces. Yet many agent benchmarks

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Code with Claude, our developer conference, returns next week. Whether you're just getting started with Claude Code or you've been building …

DGX agent

Code with Claude, our developer conference, returns next week. Whether you're just getting started with Claude Code or you've been building for a while, there's a session for you. Register for the liv

model-releasesthariq--x
1 May 2026
Model Releases

Cofounder 2 launches on May 4th.

DGX agent

Yohei Nakajima announced the launch of Cofounder 2 on May 4th via X. Cofounder is an AI agent tool designed to assist with business and startup tasks. The announcement was shared on social media to in

model-releasesyohei-nakajima--x
1 May 2026
Model Releases

COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts

DGX agent

arXiv:2604.27389v1 Announce Type: cross Abstract: In recent years, Multimodal Large Language Models (MLLMs) have achieved remarkable progress on a wide range of multimodal benchmarks. Despite these ad

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Contextual Agentic Memory is a Memo, Not True Memory

DGX agent

arXiv:2604.27707v1 Announce Type: new Abstract: Current agentic memory systems (vector stores, retrieval-augmented generation, scratchpads, and context-window management) do not implement memory: they

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Cross-Lingual Response Consistency in Large Language Models: An ILR-Informed Evaluation of Claude Across Six Languages

DGX agent

arXiv:2604.27137v1 Announce Type: new Abstract: This paper introduces a systematic evaluation framework grounded in the Interagency Language Roundtable (ILR) Skill Level Descriptions and applies it to

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Curious about Codex? It's time to switch. You can migrate to Codex directly in the Codex app and the CLI. https://chatgpt.com/codex/switch-t…

DGX agent

OpenAI announced a migration feature allowing users to switch to Codex directly through both the Codex app and command-line interface (CLI). The announcement indicates that Codex migration tools are n

model-releasesopenai--x
1 May 2026
Model Releases

Decoding Scientific Experimental Images: The SPUR Benchmark for Perception, Understanding, and Reasoning

DGX agent

arXiv:2604.27604v1 Announce Type: new Abstract: We introduce SPUR, a comprehensive benchmark for scientific experimental image perception, understanding, and reasoning, comprising 4,264 question-answe

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

DeepTutor: Towards Agentic Personalized Tutoring

DGX agent

arXiv:2604.26962v1 Announce Type: cross Abstract: Education represents one of the most promising real-world applications for Large Language Models (LLMs). However, conventional tutoring systems rely o

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures

DGX agent

arXiv:2604.28118v1 Announce Type: cross Abstract: Transformer models are widely deployed in critical AI applications, yet faults in their attention mechanisms, projections, and other internal componen

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Diagnosing Capability Gaps in Fine-Tuning Data

DGX agent

arXiv:2604.27547v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) for domain-specific tasks requires training datasets that comprehensively cover the target capabilities a pract

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Do Papers Tell the Whole Story? A Benchmark and Framework for Uncovering Hidden Implementation Gaps in Bioinformatics

DGX agent

arXiv:2603.22018v2 Announce Type: replace Abstract: Ensuring consistency between research papers and their corresponding software code implementations is a fundamental prerequisite for guaranteeing th

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Do What I Say: A Spoken Prompt Dataset for Instruction-Following

DGX agent

arXiv:2603.09881v2 Announce Type: replace Abstract: Speech Large Language Models (SLLMs) have rapidly expanded, supporting a wide range of tasks. These models are typically evaluated using text prompt

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Do World Action Models Generalize Better than VLAs? A Robustness Study

DGX agent

arXiv:2603.22078v3 Announce Type: replace Abstract: Robot action planning in the real world is challenging as it requires not only understanding the current state of the environment but also predictin

model-releasesarxiv-cs-ro
1 May 2026
Model Releases

DPN-LE: Dual Personality Neuron Localization and Editing for Large Language Models

DGX agent

arXiv:2604.27929v1 Announce Type: new Abstract: With the widespread adoption of large language models (LLMs), understanding their personality representation mechanisms has become critical. As a novel

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Dynamic Scaled Gradient Descent for Stable Fine-Tuning for Classifications

DGX agent

arXiv:2604.27987v1 Announce Type: new Abstract: Fine-tuning pretrained models has become a standard approach to adapting pretrained knowledge to improve the accuracy on new sparse, imbalance datasets.

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

EdgeSpike: Spiking Neural Networks for Low-Power Autonomous Sensing in Edge IoT Architectures

DGX agent

arXiv:2604.27004v1 Announce Type: cross Abstract: We propose EdgeSpike, a co-designed spiking neural network (SNN) framework for autonomous low-power sensing in edge Internet of Things (IoT) architect

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

EDU-CIRCUIT-HW: Evaluating Multimodal Large Language Models on Real-World University-Level STEM Student Handwritten Solutions

DGX agent

arXiv:2602.00095v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) hold significant promise for revolutionizing traditional education and reducing teachers' workload. H

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Entropy of Ukrainian

DGX agent

arXiv:2604.27534v1 Announce Type: new Abstract: In natural language processing, the entropy of a language is a measure of its unpredictability and complexity. The first study on this subject was condu

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Event with @googlegemma next week! Hang out with the fellow builders and members of the Gemma + LM teams at LM Studio HQ in NYC. When: Monda…

DGX agent

Event with @googlegemma next week! Hang out with the fellow builders and members of the Gemma + LM teams at LM Studio HQ in NYC. When: Monday, May 11th RSVP: required, link below High likelihood of pi

model-releaseslm-studio--x
1 May 2026
Model Releases

Exploring Interaction Paradigms for LLM Agents in Scientific Visualization

DGX agent

arXiv:2604.27996v1 Announce Type: new Abstract: This paper examines how different types of large language model (LLM) agents perform on scientific visualization (SciVis) tasks, where users generate vi

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Exploring the Limits of Pruning: Task-Specific Neurons, Model Collapse, and Recovery in Task-Specific Large Language Models

DGX agent

arXiv:2604.27115v1 Announce Type: new Abstract: Neuron pruning is widely used to reduce the computational cost and parameter footprint of large language models, yet it remains unclear whether neurons

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Fake3DGS: A Benchmark for 3D Manipulation Detection in Neural Rendering

DGX agent

arXiv:2604.27590v1 Announce Type: new Abstract: Recent advances in 3D reconstruction and neural rendering,particularly 3D Gaussian Splatting, make it feasible and simple to edit 3D scenes and re-rende

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection

DGX agent

arXiv:2501.04066v2 Announce Type: replace Abstract: As a special type of multimedia data, Lithography Hotspot Detection (LHD) training often requires stronger privacy protection than conventional mult

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Fidelity, Diversity, and Privacy: A Multi-Dimensional LLM Evaluation for Clinical Data Augmentation

DGX agent

arXiv:2604.27014v1 Announce Type: new Abstract: The scarcity of high-quality annotated medical data, particularly in mental health, poses a significant bottleneck for training robust machine learning

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

FinChain: A Symbolic Benchmark for Verifiable Chain-of-Thought Financial Reasoning

DGX agent

arXiv:2506.02515v4 Announce Type: replace-cross Abstract: Multi-step symbolic reasoning is essential for robust financial analysis; yet, current benchmarks largely overlook this capability. Existing d

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting

DGX agent

arXiv:2604.27974v1 Announce Type: new Abstract: Despite the rapid progress of large vision-language models (LVLMs), fine-grained, state-conditioned GUI interaction remains challenging. Current evaluat

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

DGX agent

arXiv:2506.07180v3 Announce Type: replace-cross Abstract: As video large language models (Video-LLMs) become increasingly integrated into real-world applications that demand grounded multimodal reason

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving

DGX agent

arXiv:2604.02715v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models have become a dominant paradigm for scaling large language models, but their rapidly growing parameter sizes introdu

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction

DGX agent

arXiv:2604.28115v1 Announce Type: cross Abstract: Existing learning-based occupancy prediction methods rely on large-scale 3D annotations and generalize poorly across environments. We present FreeOcc,

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

From Mirage to Grounding: Towards Reliable Multimodal Circuit-to-Verilog Code Generation

DGX agent

arXiv:2604.27969v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate visual artifacts into code, from UI mockups into HTML to scientific plots

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Test-taking to Cognitive Scaffolding: A Pedagogical Diagnostic Benchmark for LLMs on English Standardized Tests

DGX agent

arXiv:2505.17056v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly integrated into educational tools, current evaluations on standardized tests predominantly fo

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction

DGX agent

arXiv:2604.27906v1 Announce Type: new Abstract: Persistent AI memory is often reduced to a retrieval problem: store prior interactions as text, embed them, and ask the model to recover relevant contex

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Function-based Parametric Co-Design Optimization of Dexterous Hands

DGX agent

arXiv:2604.27557v1 Announce Type: new Abstract: Despite advances in dexterous hand manipulation, robotic hand design is still largely decoupled from task-driven evaluation and control, limiting system

model-releasesarxiv-cs-ro
1 May 2026
Model Releases

Gait Recognition via Deep Residual Networks and Multi-Branch Feature Fusion

DGX agent

arXiv:2604.27353v1 Announce Type: new Abstract: Gait recognition has emerged as a compelling biometric modality for surveillance and security applications, offering inherent advantages such as non-int

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Generalizable Sparse-View 3D Reconstruction from Unconstrained Images

DGX agent

arXiv:2604.28193v1 Announce Type: new Abstract: Reconstructing 3D scenes from sparse, unposed images remains challenging under real-world conditions with varying illumination and transient occlusions.

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Generalizing the Geometry of Model Merging Through Frechet Averages

DGX agent

arXiv:2604.27155v1 Announce Type: new Abstract: Model merging aims to combine multiple models into one without additional training. Naive parameter-space averaging can be fragile under architectural s

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Generate Your Talking Avatar from Video Reference

DGX agent

arXiv:2604.27918v1 Announce Type: new Abstract: Existing talking avatar methods typically adopt an image-to-video pipeline conditioned on a static reference image within the same scene as the target g

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Global Optimality for Constrained Exploration via Penalty Regularization

DGX agent

arXiv:2604.28144v1 Announce Type: new Abstract: Efficient exploration is a central problem in reinforcement learning and is often formalized as maximizing the entropy of the state-action occupancy mea

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

GlowQ: Group-Shared LOw-Rank Approximation for Quantized LLMs

DGX agent

arXiv:2603.25385v2 Announce Type: replace-cross Abstract: Quantization techniques such as BitsAndBytes, AWQ, and GPTQ are widely used as a standard method in deploying large language models but often

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

@GoogleAIStudio And this sample project was created on Canvas in @GeminiApp. It’s a high-speed rhythm game where you tap to the beat and col…

DGX agent

@GoogleAIStudio And this sample project was created on Canvas in @GeminiApp. It’s a high-speed rhythm game where you tap to the beat and collect power-ups to remix the track. Watch as the numbers appe

model-releasesgoogle-ai--x
1 May 2026
Model Releases

@GoogleAIStudio @GeminiApp We can’t wait to see where your creativity takes you. Vibe code your countdown idea in @GoogleAIStudio or Canvas …

DGX agent

@GoogleAIStudio @GeminiApp We can’t wait to see where your creativity takes you. Vibe code your countdown idea in @GoogleAIStudio or Canvas in @GeminiApp, then submit it here: http://goo.gle/codetheco

model-releasesgoogle-ai--x
1 May 2026
← Previous
1…367368369370371…470
Next →