AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,512 results
29 Apr 2026

Benchmarking OCR Pipelines with Adaptive Enhancement for Multi-Domain Retail Bill Digitization

Model ReleasesDGX agent

arXiv:2604.25176v1 Announce Type: new Abstract: The digitization of multi-domain retail billing documents remains a challenging task due to variability in scan quality, layout heterogeneity, and domai

Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal

Model ReleasesDGX agent

arXiv:2509.09708v3 Announce Type: replace Abstract: Refusal on harmful prompts is a key safety behaviour in instruction-tuned large language models (LLMs), yet the internal causes of this behaviour re

BifDet: A 3D Bifurcation Detection Dataset for Airway-Tree Modeling

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.24999v1 Announce Type: new Abstract: Thoracic Computed Tomography (CT) scans offer detailed insights into the intricate branching network of the airway tree, which is essential for understa

BLASST: Dynamic BLocked Attention Sparsity via Softmax Thresholding

Model ReleasesDGX agent

arXiv:2512.12087v3 Announce Type: replace Abstract: The growing demand for long-context inference capabilities in Large Language Models (LLMs) has intensified the computational and memory bottlenecks

Building the compute infrastructure for the Intelligence Age

Model ReleasesDGX agent

OpenAI outlines the computational infrastructure requirements and strategies necessary to support advanced AI systems in the emerging Intelligence Age. The piece likely discusses scaling challenges, h

CAN-QA: A Question-Answering Benchmark for Reasoning over In-Vehicle CAN Traffic

Model ReleasesDGX agent

arXiv:2604.24935v1 Announce Type: cross Abstract: The Controller Area Network (CAN) is a safety-critical in-vehicle communication protocol that lacks built-in security mechanisms, making intrusion det

CGU-ILALab at FoodBench-QA 2026: Comparing Traditional and LLM-based Approaches for Recipe Nutrient Estimation

Model ReleasesDGX agent

arXiv:2604.25774v1 Announce Type: new Abstract: Accurate nutrient estimation from unstructured recipe text is an important yet challenging problem in dietary monitoring, due to ambiguous ingredient te

Cheaper, Better, Faster, Stronger: Robust Text-to-SQL without Chain-of-Thought or Fine-Tuning

Model ReleasesDGX agent

arXiv:2505.14174v2 Announce Type: replace Abstract: LLMs are effective at code generation tasks like text-to-SQL, but is it worth the cost? Many state-of-the-art approaches use non-task-specific LLM t

Citation Failure: Definition, Analysis and Efficient Mitigation

Model ReleasesDGX agent

arXiv:2510.20303v3 Announce Type: replace Abstract: Citations from LLM-based RAG systems are supposed to simplify response verification. However, this goal is undermined in cases of citation failure,

Codex can help you compare choices against your criteria and keep track of the tradeoffs.

Model ReleasesDGX agent

Codex is an OpenAI tool that assists users in evaluating multiple options by comparing them against specified criteria and documenting the associated tradeoffs. This capability helps users make more i

Combating Visual Neglect and Semantic Drift in Large Multimodal Models for Enhanced Cross-Modal Retrieval

Model ReleasesDGX agent

arXiv:2604.25273v1 Announce Type: new Abstract: Despite significant progress in Unified Multimodal Retrieval (UMR) powered by Large Multimodal Models (LMMs), existing embedding methods primarily focus

Command Zero opens its autonomous security operations center platform with APIs and an MCP server

Model ReleasesDGX agent

Cyber investigations platform provider Command Zero Inc. today released a set of application programming interface endpoints and a Model Context Protocol server for its autonomous security operations

Comparing Data Assimilation and Likelihood-Based Inference on Latent State Estimation in Agent-Based Models

Model ReleasesDGX agent

arXiv:2509.17625v2 Announce Type: replace Abstract: In this paper, we present the first systematic comparison of Data Assimilation (DA) and Likelihood-Based Inference (LBI) in the context of an Agent-

Contrast-Enhanced Gating in GRUs for Robust Low-Data Sequence Learning

Model ReleasesDGX agent

arXiv:2402.09034v3 Announce Type: replace Abstract: Activation functions govern how recurrent networks regulate and transmit information across temporal dependencies. Despite advances in sequence mode

CoRE: Concept-Reasoning Expansion for Continual Brain Lesion Segmentation

Model ReleasesDGX agent

arXiv:2604.25376v1 Announce Type: new Abstract: Accurate brain lesion segmentation in MRI is vital for effective clinical diagnosis and treatment planning. Due to high annotation costs and strict data

CRAFT: Grounded Multi-Agent Coordination Under Partial Information

Model ReleasesDGX agent

arXiv:2603.25268v2 Announce Type: replace Abstract: We introduce CRAFT, a multi-agent benchmark for evaluating pragmatic communication in large language models under strict partial information. In thi

Cross-Lingual Jailbreak Detection via Semantic Codebooks

Model ReleasesDGX agent

arXiv:2604.25716v1 Announce Type: new Abstract: Safety mechanisms for large language models (LLMs) remain predominantly English-centric, creating systematic vulnerabilities in multilingual deployment.

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation

Model ReleasesDGX agent

arXiv:2604.25318v1 Announce Type: cross Abstract: Cutscenes are carefully choreographed cinematic sequences embedded in video games and interactive media, serving as the primary vehicle for narrative

Cybersecurity in the Intelligence Age

Model ReleasesDGX agent

This OpenAI resource discusses cybersecurity challenges and strategies in an era dominated by artificial intelligence and advanced threat actors. It likely covers how AI technologies are transforming

Data-Driven Hamiltonian Reduction for Superconducting Qubits via Meta-Learning

Model ReleasesDGX agent

arXiv:2604.24912v1 Announce Type: cross Abstract: We introduce HAML (Hamiltonian Adaptation via Meta-Learning), a framework for fast online adaptation of effective Hamiltonian models of superconductin

Day 2 at AI Dev 26 in SF surrounded by the best builders. 📍Find us at booth 121. @DeepLearningAI

Model ReleasesDGX agent

AI21 Labs is exhibiting at AI Dev 26 conference in San Francisco, located at booth 121, highlighting their participation among leading AI developers and builders in the industry. The post indicates th

DeepSeek has began grayscale testing for DeepSeek with Vision

Model ReleasesDGX agent

DeepSeek V4 is undergoing limited grayscale testing with a new interface featuring Fast, Expert, and Vision modes . The Vision version represents the multimodal component of the upcoming DeepSeek V4 r

DeepSeek V4 Pro is crazy good at bug fixing. Costs counted in cents not dollars/tens of dollars. It’s the next level quiet confidence not to…

Model ReleasesDGX agent

DeepSeek V4 Pro is crazy good at bug fixing. Costs counted in cents not dollars/tens of dollars. It’s the next level quiet confidence not to get caught up by benchmarks & rankings, and just let us use

DeepSeek-V4 Pro now available on Together AI

Model ReleasesDGX agent

DeepSeek-V4 Pro is now available on Together AI with 512K context, controllable reasoning modes, and cached-input pricing for long-context reasoning workloads like code agents, document intelligence,

Despite constant chants of “exponential progress”, trust issues continue to plague generative AI.

Model ReleasesDGX agent

Despite constant chants of “exponential progress”, trust issues continue to plague generative AI. OPUS 4.7 JUST MASS EMAILED AN ENTIRE DATABASE 20 TIMES PER CONTACT. WITHOUT PERMISSION a developer had

Detecting Dental Landmarks from Intraoral 3D Scans: the 3DTeethLand challenge

Model ReleasesDGX agent

arXiv:2512.08323v2 Announce Type: replace Abstract: Teeth landmark detection is a key task in modern orthodontics, supporting advanced diagnosis, personalized treatment planning, and effective monitor

DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA

Model ReleasesDGX agent

arXiv:2603.29844v2 Announce Type: replace-cross Abstract: The development of Vision-Language-Action (VLA) models has been significantly accelerated by pre-trained Vision-Language Models (VLMs). Howeve

Dictionary learning for Kernel EDMD

Model ReleasesDGX agent

arXiv:2604.25572v1 Announce Type: cross Abstract: Studying nonlinear dynamical systems through their state space behavior can be challenging, and one possible alternative is to analyze them via their

🤖👇 Did you know that you can use the @GoogleAIStudio Gemini Live API for robotics?! Check out @ThePracticalDev blog post from my colleague…

Model ReleasesDGX agent

🤖👇 Did you know that you can use the @GoogleAIStudio Gemini Live API for robotics?! Check out @ThePracticalDev blog post from my colleague @thorwebdev to learn more: Finally added Gemini Live to @poll

DiRe-RAPIDS: Topology-faithful dimensionality reduction at scale

Model ReleasesDGX agent

arXiv:2604.25209v1 Announce Type: new Abstract: Dimensionality reduction methods such as UMAP and t-SNE are central tools for visualising high-dimensional data, but their local-neighborhood objectives

Divine, a Vine reboot financed by Jack Dorsey-backed nonprofit 'and Other Stuff' and built by an early Twitter employee, debuts with ~500K restored Vine videos (Sarah Perez/TechCrunch)

Model ReleasesDGX agent

Sarah Perez / TechCrunch: Divine, a Vine reboot financed by Jack Dorsey-backed nonprofit “and Other Stuff” and built by an early Twitter employee, debuts with ~500K restored Vine videos — A new projec

Doing More With Less: Revisiting the Effectiveness of LLM Pruning for Test-Time Scaling

Model ReleasesDGX agent

arXiv:2604.25098v1 Announce Type: cross Abstract: While current Large Language Models (LLMs) exhibit remarkable reasoning capabilities through test-time compute scaling (TTS), their massive parameter

Dont Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination

Model ReleasesDGX agent

arXiv:2604.24978v1 Announce Type: new Abstract: Enterprise deep research often fails to produce decision-ready reports due to uneven information coverage, context explosion, and premature stopping. We

DRAGON: A Benchmark for Evidence-Grounded Visual Reasoning over Diagrams

Model ReleasesDGX agent

arXiv:2604.25231v1 Announce Type: cross Abstract: Diagram question answering (DQA) requires models to interpret structured visual representations such as charts, maps, infographics, circuit schematics

DV-World: Benchmarking Data Visualization Agents in Real-World Scenarios

Model ReleasesDGX agent

arXiv:2604.25914v1 Announce Type: new Abstract: Real-world data visualization (DV) requires native environmental grounding, cross-platform evolution, and proactive intent alignment. Yet, existing benc

Egocentric Tactile and Proximity Sensors as Observation Priors for Humanoid Collision Avoidance

Model ReleasesDGX agent

arXiv:2604.25554v1 Announce Type: cross Abstract: Collision-free motion is often aided by tactile and proximity sensors distributed on the body of the robot due to their resistance to occlusion as opp

Elite-Driven Support Vector Machines for Classification

Model ReleasesDGX agent

arXiv:2604.25158v1 Announce Type: cross Abstract: Support vector machines (SVMs) are a standard tool for binary classification, but their classical formulations are purely data-driven and offer no dir

Enhancing Financial Report Question-Answering: A Retrieval-Augmented Generation System with Reranking Analysis

Model ReleasesDGX agent

arXiv:2603.16877v2 Announce Type: replace Abstract: Financial analysts face significant challenges extracting information from lengthy 10-K reports, which often exceed 100 pages. This paper presents a

EOS-Bench: A Comprehensive Benchmark for Earth Observation Satellite Scheduling

Model ReleasesDGX agent

arXiv:2604.25782v1 Announce Type: cross Abstract: Earth observation satellite imaging scheduling is a challenging NP-hard combinatorial optimisation problem central to space mission operations. While

ESICA: A Scalable Framework for Text-Guided 3D Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.24876v1 Announce Type: new Abstract: Text guided 3D medical image segmentation offers a flexible alternative to class based and spatial prompt based models by allowing users to specify regi

Evaluating Computational Pathology Foundation Models for Prostate Cancer Grading under Distribution Shifts

Model ReleasesDGX agent

arXiv:2410.06723v2 Announce Type: replace-cross Abstract: Pathology foundation models (PFMs) have emerged as powerful pretrained encoders for computational pathology, but their robustness under clinic

Evaluating LLM Safety Under Repeated Inference via Accelerated Prompt Stress Testing

Model ReleasesDGX agent

arXiv:2602.11786v2 Announce Type: replace Abstract: Traditional benchmarks for large language models (LLMs), such as HELM and AIR-BENCH, primarily assess safety through breadth-oriented evaluation acr

EvoTSC: Evolving Feature Learning Models for Time Series Classification via Genetic Programming

Model ReleasesDGX agent

arXiv:2604.25499v1 Announce Type: new Abstract: Time series classification is an important analytical task across diverse domains. However, its practical application is often hindered by the scarcity

Exploring Reasoning Reward Model for Agents

Model ReleasesDGX agent

arXiv:2601.22154v2 Announce Type: replace-cross Abstract: Agentic Reinforcement Learning (Agentic RL) has achieved notable success in enabling agents to perform complex reasoning and tool use. However

Exploring Remote Photoplethysmography for Neonatal Pain Detection from Facial Videos

Model ReleasesDGX agent

arXiv:2604.25680v1 Announce Type: new Abstract: Unaddressed pain in neonates can lead to adverse effects, including delayed development and slower weight gain, emphasising the need for more objective

Faithful Autoformalization via Roundtrip Verification and Repair

Model ReleasesDGX agent

arXiv:2604.25031v1 Announce Type: new Abstract: When an LLM formalizes natural language, how do we know the output is faithful? We propose a roundtrip verification approach which does not require grou

Faithfulness-QA: A Counterfactual Entity Substitution Dataset for Training Context-Faithful RAG Models

Model ReleasesDGX agent

arXiv:2604.25313v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) models frequently produce answers grounded in parametric memory rather than the retrieved context, undermining the

Falcon Heavy

Model ReleasesDGX agent

Falcon Heavy All systems are looking good for Falcon Heavy’s launch of the @viasat-3 F3 mission from Florida. The 85-minute window opens at 10:13 a.m. ET. Teams are keeping an eye on weather → http://

Falcon Heavy launches the @viasat-3 F3 mission to orbit from Florida

Model ReleasesDGX agent

SpaceX's Falcon Heavy rocket successfully launched the Viasat-3 F3 satellite mission from Florida, placing the communications satellite into orbit. Viasat-3 is part of a series of high-capacity satell

FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments

Model ReleasesDGX agent

arXiv:2604.25135v1 Announce Type: new Abstract: Large Language Models are being increasingly deployed as the decision-making core of autonomous agents capable of effecting change in external environme

FARM: Enhancing Molecular Representations with Functional Group Awareness

Model ReleasesDGX agent

arXiv:2410.02082v4 Announce Type: replace Abstract: We introduce Functional Group-Aware Representations for Small Molecules (FARM), a novel foundation model designed to bridge the gap between SMILES,

FCMBench-Video: Benchmarking Document Video Intelligence

Model ReleasesDGX agent

arXiv:2604.25186v1 Announce Type: new Abstract: Document understanding is a critical capability in financial credit review, onboarding, and remote verification, where both decision accuracy and eviden

Feasible-First Exploration for Constrained ML Deployment Optimization in Crash-Prone Hierarchical Search Spaces

Model ReleasesDGX agent

arXiv:2604.25073v1 Announce Type: new Abstract: Deploying machine learning models under production constraints requires joint optimization over model family, quantization scheme, runtime backend, and

FED-FSTQ: Fisher-Guided Token Quantization for Communication-Efficient Federated Fine-Tuning of LLMs on Edge Devices

Model ReleasesDGX agent

arXiv:2604.25421v1 Announce Type: new Abstract: Federated fine-tuning provides a practical route to adapt large language models (LLMs) on edge devices without centralizing private data, yet in mobile

Forward and backward benchmark results across common configurations.

Model ReleasesDGX agent

The document titled 'Forward and backward benchmark results across common configurations' is inaccessible because JavaScript is disabled in the browser. Users are prompted to enable JavaScript or swit

From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models

Model ReleasesDGX agent

arXiv:2510.18030v2 Announce Type: replace Abstract: Structured pruning is a practical approach to deploying large language models (LLMs) efficiently, as it yields compact, hardware-friendly architectu

Frontier Coding Agents Can Now Implement an AlphaZero Self-Play Machine Learning Pipeline For Connect Four That Performs Comparably to an External Solver

Model ReleasesDGX agent

arXiv:2604.25067v1 Announce Type: cross Abstract: Forecasting when AI systems will become capable of meaningfully accelerating AI research is a central challenge for AI safety. Existing benchmarks mea

G-Loss: Graph-Guided Fine-Tuning of Language Models

Model ReleasesDGX agent

arXiv:2604.25853v1 Announce Type: new Abstract: Traditional loss functions, including cross-entropy, contrastive, triplet, and su pervised contrastive losses, used for fine-tuning pre-trained language

GAIA-v2-LILT: Multilingual Adaptation of Agent Benchmark beyond Translation

Model ReleasesDGX agent

arXiv:2604.24929v1 Announce Type: new Abstract: Agent benchmarks remain largely English-centric, while their multilingual versions are often built with machine translation (MT) and limited post-editin

Gemini now can create documents, and it is a nice start, but not up to the frontier yet, as you can see from my 'LBO of Hogwarts' test. Powe…

Model ReleasesDGX agent

Gemini now can create documents, and it is a nice start, but not up to the frontier yet, as you can see from my 'LBO of Hogwarts' test. PowerPoints are substantially worse than NotebookLM, spreadsheet

← Previous
1…299300301302303…376
Next →