AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,555 results
Model Releases

Can Large Language Models Implement Agent-Based Models? An ODD-based Replication Study

DGX agent

arXiv:2602.10140v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can now synthesize non-trivial executable code from textual descriptions, raising an important question: can LLMs

model-releasesarxiv-cs-ai
1 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

CareGuardAI: Context-Aware Multi-Agent Guardrails for Clinical Safety & Hallucination Mitigation in Patient-Facing LLMs

DGX agent

arXiv:2604.26959v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into patient-facing healthcare systems offers significant potential to improve access to medical information.

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

CausalCompass: Evaluating the Robustness of Time-Series Causal Discovery in Misspecified Scenarios

DGX agent

arXiv:2602.07915v2 Announce Type: replace-cross Abstract: Causal discovery from time series is a fundamental task in machine learning. However, its widespread adoption is hindered by a reliance on unt

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Characterizing the Consistency of the Emergent Misalignment Persona

DGX agent

arXiv:2604.28082v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) on narrowly misaligned data generalizes to broadly misaligned behavior, a phenomenon termed emergent misalignme

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ChipLingo: A Systematic Training Framework for Large Language Models in EDA

DGX agent

arXiv:2604.27415v1 Announce Type: new Abstract: With the rapid advancement of semiconductor technology, Electronic Design Automation (EDA) has become an increasingly knowledge-intensive and document-d

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

CL-bench Life: Can Language Models Learn from Real-Life Context?

DGX agent

arXiv:2604.27043v1 Announce Type: new Abstract: Today's AI assistants such as OpenClaw are designed to handle context effectively, making context learning an increasingly important capability for mode

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows

DGX agent

arXiv:2604.28139v1 Announce Type: cross Abstract: LLM agents are expected to complete end-to-end units of work across software tools, business services, and local workspaces. Yet many agent benchmarks

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Code with Claude, our developer conference, returns next week. Whether you're just getting started with Claude Code or you've been building …

DGX agent

Code with Claude, our developer conference, returns next week. Whether you're just getting started with Claude Code or you've been building for a while, there's a session for you. Register for the liv

model-releasesthariq--x
1 May 2026
Model Releases

Cofounder 2 launches on May 4th.

DGX agent

Yohei Nakajima announced the launch of Cofounder 2 on May 4th via X. Cofounder is an AI agent tool designed to assist with business and startup tasks. The announcement was shared on social media to in

model-releasesyohei-nakajima--x
1 May 2026
Model Releases

COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts

DGX agent

arXiv:2604.27389v1 Announce Type: cross Abstract: In recent years, Multimodal Large Language Models (MLLMs) have achieved remarkable progress on a wide range of multimodal benchmarks. Despite these ad

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Contextual Agentic Memory is a Memo, Not True Memory

DGX agent

arXiv:2604.27707v1 Announce Type: new Abstract: Current agentic memory systems (vector stores, retrieval-augmented generation, scratchpads, and context-window management) do not implement memory: they

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Cross-Lingual Response Consistency in Large Language Models: An ILR-Informed Evaluation of Claude Across Six Languages

DGX agent

arXiv:2604.27137v1 Announce Type: new Abstract: This paper introduces a systematic evaluation framework grounded in the Interagency Language Roundtable (ILR) Skill Level Descriptions and applies it to

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Curious about Codex? It's time to switch. You can migrate to Codex directly in the Codex app and the CLI. https://chatgpt.com/codex/switch-t…

DGX agent

OpenAI announced a migration feature allowing users to switch to Codex directly through both the Codex app and command-line interface (CLI). The announcement indicates that Codex migration tools are n

model-releasesopenai--x
1 May 2026
Model Releases

Decoding Scientific Experimental Images: The SPUR Benchmark for Perception, Understanding, and Reasoning

DGX agent

arXiv:2604.27604v1 Announce Type: new Abstract: We introduce SPUR, a comprehensive benchmark for scientific experimental image perception, understanding, and reasoning, comprising 4,264 question-answe

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

DeepTutor: Towards Agentic Personalized Tutoring

DGX agent

arXiv:2604.26962v1 Announce Type: cross Abstract: Education represents one of the most promising real-world applications for Large Language Models (LLMs). However, conventional tutoring systems rely o

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures

DGX agent

arXiv:2604.28118v1 Announce Type: cross Abstract: Transformer models are widely deployed in critical AI applications, yet faults in their attention mechanisms, projections, and other internal componen

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Diagnosing Capability Gaps in Fine-Tuning Data

DGX agent

arXiv:2604.27547v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) for domain-specific tasks requires training datasets that comprehensively cover the target capabilities a pract

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Do Papers Tell the Whole Story? A Benchmark and Framework for Uncovering Hidden Implementation Gaps in Bioinformatics

DGX agent

arXiv:2603.22018v2 Announce Type: replace Abstract: Ensuring consistency between research papers and their corresponding software code implementations is a fundamental prerequisite for guaranteeing th

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Do What I Say: A Spoken Prompt Dataset for Instruction-Following

DGX agent

arXiv:2603.09881v2 Announce Type: replace Abstract: Speech Large Language Models (SLLMs) have rapidly expanded, supporting a wide range of tasks. These models are typically evaluated using text prompt

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Do World Action Models Generalize Better than VLAs? A Robustness Study

DGX agent

arXiv:2603.22078v3 Announce Type: replace Abstract: Robot action planning in the real world is challenging as it requires not only understanding the current state of the environment but also predictin

model-releasesarxiv-cs-ro
1 May 2026
Model Releases

DPN-LE: Dual Personality Neuron Localization and Editing for Large Language Models

DGX agent

arXiv:2604.27929v1 Announce Type: new Abstract: With the widespread adoption of large language models (LLMs), understanding their personality representation mechanisms has become critical. As a novel

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Dynamic Scaled Gradient Descent for Stable Fine-Tuning for Classifications

DGX agent

arXiv:2604.27987v1 Announce Type: new Abstract: Fine-tuning pretrained models has become a standard approach to adapting pretrained knowledge to improve the accuracy on new sparse, imbalance datasets.

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

EdgeSpike: Spiking Neural Networks for Low-Power Autonomous Sensing in Edge IoT Architectures

DGX agent

arXiv:2604.27004v1 Announce Type: cross Abstract: We propose EdgeSpike, a co-designed spiking neural network (SNN) framework for autonomous low-power sensing in edge Internet of Things (IoT) architect

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

EDU-CIRCUIT-HW: Evaluating Multimodal Large Language Models on Real-World University-Level STEM Student Handwritten Solutions

DGX agent

arXiv:2602.00095v3 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) hold significant promise for revolutionizing traditional education and reducing teachers' workload. H

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Entropy of Ukrainian

DGX agent

arXiv:2604.27534v1 Announce Type: new Abstract: In natural language processing, the entropy of a language is a measure of its unpredictability and complexity. The first study on this subject was condu

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Event with @googlegemma next week! Hang out with the fellow builders and members of the Gemma + LM teams at LM Studio HQ in NYC. When: Monda…

DGX agent

Event with @googlegemma next week! Hang out with the fellow builders and members of the Gemma + LM teams at LM Studio HQ in NYC. When: Monday, May 11th RSVP: required, link below High likelihood of pi

model-releaseslm-studio--x
1 May 2026
Model Releases

Exploring Interaction Paradigms for LLM Agents in Scientific Visualization

DGX agent

arXiv:2604.27996v1 Announce Type: new Abstract: This paper examines how different types of large language model (LLM) agents perform on scientific visualization (SciVis) tasks, where users generate vi

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Exploring the Limits of Pruning: Task-Specific Neurons, Model Collapse, and Recovery in Task-Specific Large Language Models

DGX agent

arXiv:2604.27115v1 Announce Type: new Abstract: Neuron pruning is widely used to reduce the computational cost and parameter footprint of large language models, yet it remains unclear whether neurons

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Fake3DGS: A Benchmark for 3D Manipulation Detection in Neural Rendering

DGX agent

arXiv:2604.27590v1 Announce Type: new Abstract: Recent advances in 3D reconstruction and neural rendering,particularly 3D Gaussian Splatting, make it feasible and simple to edit 3D scenes and re-rende

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection

DGX agent

arXiv:2501.04066v2 Announce Type: replace Abstract: As a special type of multimedia data, Lithography Hotspot Detection (LHD) training often requires stronger privacy protection than conventional mult

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Fidelity, Diversity, and Privacy: A Multi-Dimensional LLM Evaluation for Clinical Data Augmentation

DGX agent

arXiv:2604.27014v1 Announce Type: new Abstract: The scarcity of high-quality annotated medical data, particularly in mental health, poses a significant bottleneck for training robust machine learning

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

FinChain: A Symbolic Benchmark for Verifiable Chain-of-Thought Financial Reasoning

DGX agent

arXiv:2506.02515v4 Announce Type: replace-cross Abstract: Multi-step symbolic reasoning is essential for robust financial analysis; yet, current benchmarks largely overlook this capability. Existing d

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting

DGX agent

arXiv:2604.27974v1 Announce Type: new Abstract: Despite the rapid progress of large vision-language models (LVLMs), fine-grained, state-conditioned GUI interaction remains challenging. Current evaluat

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

DGX agent

arXiv:2506.07180v3 Announce Type: replace-cross Abstract: As video large language models (Video-LLMs) become increasingly integrated into real-world applications that demand grounded multimodal reason

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving

DGX agent

arXiv:2604.02715v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models have become a dominant paradigm for scaling large language models, but their rapidly growing parameter sizes introdu

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction

DGX agent

arXiv:2604.28115v1 Announce Type: cross Abstract: Existing learning-based occupancy prediction methods rely on large-scale 3D annotations and generalize poorly across environments. We present FreeOcc,

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

From Mirage to Grounding: Towards Reliable Multimodal Circuit-to-Verilog Code Generation

DGX agent

arXiv:2604.27969v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate visual artifacts into code, from UI mockups into HTML to scientific plots

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Test-taking to Cognitive Scaffolding: A Pedagogical Diagnostic Benchmark for LLMs on English Standardized Tests

DGX agent

arXiv:2505.17056v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly integrated into educational tools, current evaluations on standardized tests predominantly fo

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction

DGX agent

arXiv:2604.27906v1 Announce Type: new Abstract: Persistent AI memory is often reduced to a retrieval problem: store prior interactions as text, embed them, and ask the model to recover relevant contex

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Function-based Parametric Co-Design Optimization of Dexterous Hands

DGX agent

arXiv:2604.27557v1 Announce Type: new Abstract: Despite advances in dexterous hand manipulation, robotic hand design is still largely decoupled from task-driven evaluation and control, limiting system

model-releasesarxiv-cs-ro
1 May 2026
Model Releases

Gait Recognition via Deep Residual Networks and Multi-Branch Feature Fusion

DGX agent

arXiv:2604.27353v1 Announce Type: new Abstract: Gait recognition has emerged as a compelling biometric modality for surveillance and security applications, offering inherent advantages such as non-int

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Generalizable Sparse-View 3D Reconstruction from Unconstrained Images

DGX agent

arXiv:2604.28193v1 Announce Type: new Abstract: Reconstructing 3D scenes from sparse, unposed images remains challenging under real-world conditions with varying illumination and transient occlusions.

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Generalizing the Geometry of Model Merging Through Frechet Averages

DGX agent

arXiv:2604.27155v1 Announce Type: new Abstract: Model merging aims to combine multiple models into one without additional training. Naive parameter-space averaging can be fragile under architectural s

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Generate Your Talking Avatar from Video Reference

DGX agent

arXiv:2604.27918v1 Announce Type: new Abstract: Existing talking avatar methods typically adopt an image-to-video pipeline conditioned on a static reference image within the same scene as the target g

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Global Optimality for Constrained Exploration via Penalty Regularization

DGX agent

arXiv:2604.28144v1 Announce Type: new Abstract: Efficient exploration is a central problem in reinforcement learning and is often formalized as maximizing the entropy of the state-action occupancy mea

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

GlowQ: Group-Shared LOw-Rank Approximation for Quantized LLMs

DGX agent

arXiv:2603.25385v2 Announce Type: replace-cross Abstract: Quantization techniques such as BitsAndBytes, AWQ, and GPTQ are widely used as a standard method in deploying large language models but often

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

@GoogleAIStudio And this sample project was created on Canvas in @GeminiApp. It’s a high-speed rhythm game where you tap to the beat and col…

DGX agent

@GoogleAIStudio And this sample project was created on Canvas in @GeminiApp. It’s a high-speed rhythm game where you tap to the beat and collect power-ups to remix the track. Watch as the numbers appe

model-releasesgoogle-ai--x
1 May 2026
Model Releases

@GoogleAIStudio @GeminiApp We can’t wait to see where your creativity takes you. Vibe code your countdown idea in @GoogleAIStudio or Canvas …

DGX agent

@GoogleAIStudio @GeminiApp We can’t wait to see where your creativity takes you. Vibe code your countdown idea in @GoogleAIStudio or Canvas in @GeminiApp, then submit it here: http://goo.gle/codetheco

model-releasesgoogle-ai--x
1 May 2026
← Previous
1…367368369370371…470
Next →