AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
10 Apr 2026

ClawLess: A Security Model of AI Agents

AgentsDGX agent

arXiv:2604.06284v1 Announce Type: cross Abstract: Autonomous AI agents powered by Large Language Models can reason, plan, and execute complex tasks, but their ability to autonomously retrieve informat

Daily and Weekly Periodicity in Large Language Model Performance and Its Implications for Research

ResearchDGX agent

arXiv:2602.15889v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used in research as both tools and objects of study. Much of this work assumes that LLM performa

Diffusion Language Models Know the Answer Before Decoding

ResearchDGX agent

arXiv:2508.19982v5 Announce Type: replace Abstract: Diffusion language models (DLMs) have recently emerged as an alternative to autoregressive approaches, offering parallel sequence generation and fle

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Digital Skin, Digital Bias: Uncovering Tone-Based Biases in LLMs and Emoji Embeddings

Model ReleasesDGX agent

arXiv:2604.06863v1 Announce Type: cross Abstract: Skin-toned emojis are crucial for fostering personal identity and social inclusion in online communication. As AI models, particularly Large Language

Gemma 4 on AI Gateway

Model ReleasesDGX agent

Google's Gemma 4 models — the 26B (Mixture of Experts) and 31B (Dense) variants — are now available on Vercel AI Gateway. Built on the same architecture as Gemini 3, both open models support functi...

Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Models

SafetyDGX agent

arXiv:2604.06767v1 Announce Type: new Abstract: Language models operate on discrete tokens but compute in continuous vector spaces, inducing a Voronoi tessellation over the representation manifold. We

Hallucination Detection and Evaluation of Large Language Model

Local AiDGX agent

arXiv:2512.22416v2 Announce Type: replace Abstract: Hallucinations in Large Language Models (LLMs) pose a significant challenge, generating misleading or unverifiable content that undermines trust and

I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model - it feels like the AI that you c…

ToolsDGX agent

I think it's non-obvious to many people that the OpenAI voice mode runs on a much older, much weaker model - it feels like the AI that you can talk to should be the smartest AI but it really isn't Jud

Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM

SafetyDGX agent

arXiv:2604.08425v1 Announce Type: cross Abstract: When humans label subjective content, they disagree, and that disagreement is not noise. It reflects genuine differences in perspective shaped by anno

Linear Representations of Hierarchical Concepts in Language Models

TutorialsDGX agent

arXiv:2604.07886v1 Announce Type: new Abstract: We investigate how and to what extent hierarchical relations (e.g., Japan subset Eastern Asia subset Asia) are encoded in the internal representat

MO-RiskVAE: A Multi-Omics Variational Autoencoder for Survival Risk Modeling in Multiple MyelomaMO-RiskVAE

SafetyDGX agent

arXiv:2604.06267v1 Announce Type: cross Abstract: Multimodal variational autoencoders (VAEs) have emerged as a powerful framework for survival risk modeling in multiple myeloma by integrating heteroge

MorphDistill: Distilling Unified Morphological Knowledge from Pathology Foundation Models for Colorectal Cancer Survival Prediction

SafetyDGX agent

arXiv:2604.06390v1 Announce Type: cross Abstract: Background: Colorectal cancer (CRC) remains a leading cause of cancer-related mortality worldwide. Accurate survival prediction is essential for treat

OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks

SafetyDGX agent

arXiv:2604.08539v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as the de facto Reinforcement Learning (RL) objective driving recent advancements in Multimodal

Part^{2}GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting

SafetyDGX agent

arXiv:2506.17212v2 Announce Type: replace Abstract: Articulated objects are common in the real world, yet modeling their structure and motion remains a challenging task for 3D reconstruction methods.

Personalizing Text-to-Image Generation to Individual Taste

SafetyDGX agent

arXiv:2604.07427v1 Announce Type: new Abstract: Modern text-to-image (T2I) models generate high-fidelity visuals but remain indifferent to individual user preferences. While existing reward models opt

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation

Model ReleasesDGX agent

arXiv:2512.23994v3 Announce Type: replace-cross Abstract: Text-to-audio-video (T2AV) generation is central to applications such as filmmaking and world modeling. However, current models often fail to

Privacy Attacks on Image AutoRegressive Models

ResearchDGX agent

arXiv:2502.02514v5 Announce Type: replace Abstract: Image AutoRegressive generation has emerged as a new powerful paradigm with image autoregressive models (IARs) matching state-of-the-art diffusion m

Rethinking Data Mixing from the Perspective of Large Language Models

ResearchDGX agent

arXiv:2604.07963v1 Announce Type: new Abstract: Data mixing strategy is essential for large language model (LLM) training. Empirical evidence shows that inappropriate strategies can significantly redu

Robustness Risk of Conversational Retrieval: Identifying and Mitigating Noise Sensitivity in Qwen3-Embedding Model

Model ReleasesDGX agent

arXiv:2604.06176v1 Announce Type: cross Abstract: We present an empirical study of embedding-based retrieval under realistic conversational settings, where queries are short, dialogue-like, and weakly

Stacked from One: Multi-Scale Self-Injection for Context Window Extension

Model ReleasesDGX agent

arXiv:2603.04759v2 Announce Type: replace Abstract: The limited context window of contemporary large language models (LLMs) remains a primary bottleneck for their broader application across diverse do

Weight Group-wise Post-Training Quantization for Medical Foundation Model

ResearchDGX agent

arXiv:2604.07674v1 Announce Type: new Abstract: Foundation models have achieved remarkable results in medical image analysis. However, its large network architecture and high computational complexity

When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't

Model ReleasesDGX agent

arXiv:2604.06422v1 Announce Type: cross Abstract: Understanding when Vision-Language Models (VLMs) will behave unexpectedly, whether models can reliably predict their own behavior, and if models adher

9 Apr 2026

🚀Deep Agents deploy Today we’re launching Deep Agents deploy in beta. Deep Agents deploy is the fastest way to deploy a model agnostic, ope…

AgentsDGX agent

🚀Deep Agents deploy Today we’re launching Deep Agents deploy in beta. Deep Agents deploy is the fastest way to deploy a model agnostic, open source agent harness in a production ready way. Open harnes

8 Apr 2026

Did Mythos leak Claude Code

Model ReleasesDGX agent

Anthropic's unreleased frontier model, Claude Mythos (internally codenamed 'Capybara'), was exposed through two separate leaks within the same week in late March/early April 2026. The Mythos model...

16 Aug 2026

Yutori's browser-use agents run in tight loops: screenshot, action, repeat, dozens of times per task. On Together AI, their Navigator model …

ToolsDGX agent

Yutori's browser-use agents run in tight loops: screenshot, action, repeat, dozens of times per task. On Together AI, their Navigator model beats frontier performance at 2x faster inference and 4-5x l

14 Aug 2026

AaLLM: An End-to-End Analog Circuit Design Framework from Topology Generation to Sizing Using Large Language Models

AgentsDGX agent

arXiv:2608.13472v1 Announce Type: cross Abstract: Analog circuit design is a time-consuming, iterative process in a nonlinear and high-dimensional design space that relies heavily on expert intuition.

Agreement Is Not Alignment: Divergent Moral Grounds in Human and LLM Ethical Judgments

Model ReleasesDGX agent

arXiv:2608.12368v1 Announce Type: new Abstract: Agreement with human judgments is a common proxy for evaluating the alignment of large language models (LLMs). Yet agreement in final labels does not sh

CASA: Content-Acoustic Speaking Assessment with Speech Encoder and Large Language Model

ResearchDGX agent

arXiv:2608.13101v1 Announce Type: new Abstract: Research on automatic speaking assessment (ASA) has increasingly adopted multimodal speech large language models to assess learners' speaking performanc

CONFIRMED: @c_valenzuelab, cofounder and co-CEO of @runwayml, is speaking at Thesis, our inaugural conference on work and AI. Language model…

IndustryDGX agent

CONFIRMED: @c_valenzuelab, cofounder and co-CEO of @runwayml, is speaking at Thesis, our inaugural conference on work and AI. Language models predict the next token, aka the next unit of text. Runway’

Into the ORBIT for Time Series: Training Regimes for Foundation Models

SafetyDGX agent

arXiv:2608.13262v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have advanced primarily through architectural innovation, while training regimes for large-scale heterogeneous c

Open models are improving at an unprecedented rate. Congrats to the whole @Zai_org team. Ollama will support GLM-5.3 upon the open release! …

Local AiDGX agent

Open models are improving at an unprecedented rate. Congrats to the whole @Zai_org team. Ollama will support GLM-5.3 upon the open release! ❤️ Get ready to code. Introducing GLM-5.3: Built to Code. Re

Reduced Order Modeling for Tsunami Forecasting with Bayesian Hierarchical Pooling

ApplicationsDGX agent

arXiv:2512.19804v2 Announce Type: replace Abstract: Reduced-order models (ROMs) can represent spatiotemporal processes in significantly fewer dimensions and can often be solved many orders of magnitud

13 Aug 2026

Analysis of Federated Aggregation under Model Poisoning and Backdoor Attacks: A Reconstructed Cross-Dataset and Cross-Architecture Benchmark

Model ReleasesDGX agent

arXiv:2608.11423v1 Announce Type: cross Abstract: Robust comparisons of federated aggregation methods require joint consideration of predictive performance, threat definitions, metric semantics, and e

Boundary-Continuous Cross-Camera RGB Mapping via Hue-Split Model Trees

ResearchDGX agent

arXiv:2608.11548v1 Announce Type: cross Abstract: We propose a hue-split model-tree method for boundary-continuous cross-camera RGB mapping. Cross-camera RGB mapping aims to produce consistent color r

CLAIM: Leading Open-domain Active Clarification of Large Language Models with Uncertainty Measurement

SafetyDGX agent

arXiv:2608.11631v1 Announce Type: new Abstract: In open-domain human-computer interaction scenarios, large language models (LLMs) frequently encounter user queries that are ambiguous or incomplete. In

CORA-Diff: Confidence-Oriented Residual Acceptance for Efficient Diffusion Language Model Inference

ResearchDGX agent

arXiv:2608.11235v1 Announce Type: new Abstract: Diffusion language models (DLMs) update many tokens in parallel, yet practical decoders often use a fixed denoising horizon. Many predictions stabilize

Distillation of Foundation Models for Time-dependent PDEs

Local AiDGX agent

arXiv:2608.11937v1 Announce Type: new Abstract: Foundation models for time-dependent partial differential equations (PDEs) are trained on large and diverse collections of physical systems and can gene

Hamilton-Zero: A Neural Tensor-Network Foundation Model for Ground States of Arbitrary Quadratic Qubit Hamiltonians

ResearchDGX agent

arXiv:2608.11911v1 Announce Type: cross Abstract: A central promise of useful quantum advantage is the ability to compute ground states of Hamiltonian systems beyond the reach of classical simulation

I think this will turn out to be wrong, and not just because I suspect there are greater returns to more intelligent models than people expe…

AgentsDGX agent

I think this will turn out to be wrong, and not just because I suspect there are greater returns to more intelligent models than people expect Economic value comes from agents, not chatbots. And accur

Large Language Model-Driven Small-Capitalization Trading: Integrating Financial News Sentiment, Macroeconomic Indicators, and Technical Signals

ResearchDGX agent

arXiv:2608.12283v1 Announce Type: cross Abstract: Large language models can extract richer signals from financial news than fixed sentiment lexicons, and recent work has explored feeding such signals

LazyTrain: Limited-resource Allocation toward Zero-waste Yield Optimization in Large Language Model Training

SafetyDGX agent

arXiv:2608.11919v1 Announce Type: new Abstract: Training large language models on limited hardware is increasingly a scheduling problem across GPU compute, host memory, PCIe transfer, and storage band

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation

ApplicationsDGX agent

arXiv:2606.17598v2 Announce Type: replace-cross Abstract: Humans naturally leverage diverse sensing modalities to interact with the physical world, while most Vision-Language-Action (VLA) models for r

OEIS Open: How many conjectures can language models turn into theorems?

Model ReleasesDGX agent

arXiv:2608.11941v1 Announce Type: new Abstract: We construct OEIS Open, a benchmark based on 492 open mathematical conjectures from the OEIS, formalized in Lean by Tsoukalas et al. Whereas these conje

RA-ClipScore: Making Generative Model Evaluation More Interpretable

SafetyDGX agent

arXiv:2608.12088v1 Announce Type: new Abstract: Generative models can produce images nearly indistinguishable from real data, yet rigorous and interpretable evaluation remains challenging. Conventiona

RLinf-VLA: A Unified and Efficient Framework for Reinforcement Learning of Vision-Language-Action Models

ResearchDGX agent

arXiv:2510.06710v3 Announce Type: replace Abstract: Recent studies have demonstrated the potential of reinforcement learning (RL) to improve the task performance of vision-language-action (VLA) models

Sources: Demis Hassabis pitched a new independent industry AI safety entity, modeled on the IAEA, to top Trump officials before stepping down as DeepMind's CEO (Wall Street Journal)

SafetyDGX agent

Wall Street Journal: Sources: Demis Hassabis pitched a new independent industry AI safety entity, modeled on the IAEA, to top Trump officials before stepping down as DeepMind's CEO — Demis Hassabis di

TELLME: Test-Enhanced Learning for Language Model Enrichment

ResearchDGX agent

arXiv:2608.11788v1 Announce Type: cross Abstract: Continual pre-training (CPT) has been widely adopted as a method for domain adaptation in large language models. However, CPT has consistently been ac

Towards Human Motion World Models via Executable Behaviour Representations

SafetyDGX agent

arXiv:2604.18064v2 Announce Type: replace Abstract: Human motion world models should capture motion's intentionality by being executable: adaptable to different actions and capable of assessing motion

TRACES: A Benchmark for Epistemic Reliability in Scientific Reasoning by LLMs

Model ReleasesDGX agent

arXiv:2608.11415v1 Announce Type: cross Abstract: Large language models are being proposed as agents in scientific workflows, in domains where no downstream verifier exists. Such deployment assumes th

12 Aug 2026

Delays in Spiking Neural Networks: A State Space Model Approach

ResearchDGX agent

arXiv:2512.01906v3 Announce Type: replace Abstract: Spiking neural networks (SNNs) are biologically inspired, event-driven models suited for temporal data processing and energy-efficient neuromorphic

DreamOmni3: Scribble-based Editing and Generation

Model ReleasesDGX agent

arXiv:2512.22525v2 Announce Type: replace Abstract: Recently unified generation and editing models have achieved remarkable success with their impressive performance. These models rely mainly on text

Graphical Models of False Information and Fact Checking Ecosystems

ApplicationsDGX agent

arXiv:2208.11582v2 Announce Type: replace-cross Abstract: The wide spread of false information online, including misinformation and disinformation, has become a major problem for our highly digitised

Hybrid Token Compression for Vision-Language Models

ResearchDGX agent

arXiv:2512.08240v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) rely on hundreds of visual tokens, leading to high computational and memory costs. Existing compression methods

I ran DeepSeek V4 Flash 284B + DSpark on one RTX PRO 6000. The drafter was faster in RAM than VRAM.

Model ReleasesDGX agent

Hey guys, Just finished benchmarking DeepSeek V4 Flash 284B + DSpark on a single RTX PRO 6000 96GB. Short version: DSpark: ~15–17% faster generation on my coding workload On this setup, the DSpark dra

Logit Lens Supervision for Patch-Level Explanations in Vision-Language Models

ResearchDGX agent

arXiv:2602.01530v2 Announce Type: replace Abstract: Modern autoregressive Vision-Language Models (VLMs) can generate fluent answers while their visual-token representations become weakly tied to the i

MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent Manipulation

SafetyDGX agent

arXiv:2608.10166v1 Announce Type: cross Abstract: Digital watermarking has emerged as a critical technique for provenance and copyright attribution in AI-generated imagery, yet its robustness against

Multimodal Item Parameter Estimation using Simulated Response Probabilitie

Model ReleasesDGX agent

arXiv:2608.10154v1 Announce Type: cross Abstract: We present results from reconstructing multiple-choice model (MCM) and three-parameter logistic (3PL) model curves using a fine-tuned multimodal large

OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation

SafetyDGX agent

arXiv:2603.19201v3 Announce Type: replace Abstract: Contact-rich manipulation tasks, such as wiping and assembly, require accurate perception of contact forces, friction changes, and state transitions

On Understanding, Identifying, and Mitigating Vulnerabilities in Agentic Large Language Models

AgentsDGX agent

arXiv:2608.10530v1 Announce Type: cross Abstract: Large Language Models (LLMs) have undergone a shift from stateless conversational interfaces to autonomous agents capable of multi-step planning, tool

VIScore: Diagnosing Planning-Relevant Quality in Latent World Models

ResearchDGX agent

arXiv:2608.11174v1 Announce Type: new Abstract: Regulating the latent space to an isotropic Gaussian distribution provides a stable and information-maximized landscape for world model planning. Howeve

← Previous
1…145146147148149…1010
Next →