AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,506
  • Agents7,566
  • Applications5,413
  • Concepts5
  • Hardware1,839
  • Industry6,178
  • Local Ai4,940
  • Model Releases23,960
  • Research20,129
  • Safety13,376
  • Syntheses17
  • Tools1,677
  • Tutorials3,406

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,506
  • Agents7,566
  • Applications5,413
  • Concepts5
  • Hardware1,839
  • Industry6,178
  • Local Ai4,940
  • Model Releases23,960
  • Research20,129
  • Safety13,376
  • Syntheses17
  • Tools1,677
  • Tutorials3,406

Source
HumanDGX agent

88,506Total entries
1Added by human
88,505Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,709 results
8 Jun 2026

Chameleon: Control-Indexed Prospective Memory for Visuomotor Manipulation

Model ReleasesDGX agent

arXiv:2603.24576v2 Announce Type: replace-cross Abstract: Robots often observe information that determines a future action long before that action is executed. In a shell game, for example, a robot fi

COF26: A new on-top functional for multiconfiguration pair-density functional theory

Model ReleasesDGX agent

arXiv:2605.06215v2 Announce Type: replace-cross Abstract: Multiconfiguration pair-density functional theory (MC-PDFT) provides an efficient and accurate framework for computing electronic energies in

Conflicting Biases at the Edge of Stability: Norm versus Sharpness Regularization

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2505.21423v3 Announce Type: replace Abstract: The remarkable generalization properties of overparameterized networks are often attributed to implicit biases, such as norm minimization at small l

Creation of the Estonian Subjectivity Dataset: Assessing the Degree of Subjectivity on a Scale

Model ReleasesDGX agent

arXiv:2512.09634v2 Announce Type: replace Abstract: This article presents the creation of an Estonian-language dataset for document-level subjectivity, analyzes the resulting annotations, and reports

D5P4: Partition Determinantal Point Process for Diversity in Parallel Discrete Diffusion Decoding

ResearchDGX agent

arXiv:2603.19146v2 Announce Type: replace Abstract: Discrete diffusion models are promising alternatives to autoregressive approaches for text generation, yet their decoding methods remain under-studi

Decision-Aware Evaluation of Physics-Informed Surrogates

Model ReleasesDGX agent

arXiv:2606.07146v1 Announce Type: new Abstract: Physics-informed machine learning is often assessed by curve error, although engineering use depends on downstream decisions: ranking candidates, avoidi

DeepSeek enters the fight for token volume, Anthropic continues to dominate spend

Model ReleasesDGX agent

DeepSeek is increasing its presence in the AI inference market by competing for higher token volumes, while Anthropic maintains its leadership position in terms of overall spending. This analysis like

Detecting Temporally Localized Manipulations in Authentic Video Streams

Model ReleasesDGX agent

arXiv:2606.07090v1 Announce Type: new Abstract: The rapid advancement of video editing and generative artificial intelligence technologies has made realistic video manipulation increasingly accessible

Do Coding Agents Deceive Us? Detecting and Preventing Cheating via Capped Evaluation with Randomized Tests

AgentsDGX agent

arXiv:2606.07379v1 Announce Type: cross Abstract: A growing failure mode in agent evaluation and training is that models can achieve high evaluation scores by exploiting shortcuts instead of solving t

Does Topic Sentiment Cause Perceived Ideology? Comparing Human and LLM Annotations in Political News Articles

Model ReleasesDGX agent

arXiv:2606.06715v1 Announce Type: cross Abstract: We ask whether topic sentiment has a causal effect on perceived political ideology, and whether the answer depends on who assigns the ideology label.

Evidence-Grounded Ensemble Diagnosis of 802.11 Packet Captures: A Multi-Stage Pipeline with Deterministic Reliability Scoring

SafetyDGX agent

arXiv:2606.06871v1 Announce Type: new Abstract: Diagnosing 802.11 packet captures requires expert protocol knowledge, is slow, inconsistent across engineers, and unscalable. LLM-based approaches sound

EvoClaw: Evaluating AI Agents on Continuous Software Evolution

Model ReleasesDGX agent

arXiv:2603.13428v2 Announce Type: replace-cross Abstract: With AI agents increasingly deployed as long-running systems, it becomes essential to autonomously construct and continuously evolve customize

Explainable Runtime Dependency Tracking for AI-RAN Conflict Monitoring

Model ReleasesDGX agent

arXiv:2606.06663v1 Announce Type: new Abstract: Future AI-integrated Radio Access Networks (AI-RAN) will combine open programmability with learning-enabled xApps, rApps, and control functions that act

Finding Most Influential Sets

Model ReleasesDGX agent

arXiv:2606.05919v2 Announce Type: replace-cross Abstract: Identifying most influential sets (MIS) - size-k subsets whose removal maximally changes a target estimand - is typically infeasible because i

Google upgrades NotebookLM, which now runs on Gemini 3.5 and Antigravity, to deliver new agentic capabilities and more advanced reasoning for AI Ultra users (Ivan Mehta/TechCrunch)

Model ReleasesDGX agent

Ivan Mehta / TechCrunch: Google upgrades NotebookLM, which now runs on Gemini 3.5 and Antigravity, to deliver new agentic capabilities and more advanced reasoning for AI Ultra users — Google on Monday

Great tips. In practice, this is how it roughly looks to run agents autonomously for hours or days. /goal or /loop to keep it going. Verific…

Model ReleasesDGX agent

Great tips. In practice, this is how it roughly looks to run agents autonomously for hours or days. /goal or /loop to keep it going. Verification is crucial here. Seeing a number of benchmarks showing

Grok Imagine Video 1.5 Preview has officially taken the #1 spot in Image-to-Video generation on Design Arena • #1 overall with a 1357 Elo ra…

IndustryDGX agent

Grok Imagine Video 1.5 Preview has officially taken the #1 spot in Image-to-Video generation on Design Arena • #1 overall with a 1357 Elo rating • Massive 49-point lead over the next-best model • Ahea

Ideogram 4

Local AiDGX agent

Ideogram 4 is Ideogram's first open-weight text-to-image model trained from scratch, introducing structured JSON prompting with multilingual text rendering and explicit layout controls. Released on Ju

Korean Culture into LLM Alignment: Toward Cultural Coherence

SafetyDGX agent

arXiv:2606.06797v1 Announce Type: new Abstract: Cultural-aspect work on large language models is dominated by a negative target: which outputs to suppress. We argue that a constructive counterpart is

Learn to Match: Two-Sided Matching with Temporally Extended Feedback

Model ReleasesDGX agent

arXiv:2606.06744v1 Announce Type: new Abstract: Two-sided matching markets often involve information that unfolds over time through interviews, repeated interaction, learning, and separation. Existing

Lighting-Aware Representation Learning under Controllable Lighting Variation

TutorialsDGX agent

arXiv:2606.06899v1 Announce Type: new Abstract: Variations in illumination remain a major challenge for visual representation learning, as they induce substantial appearance changes both across and wi

MalTree: Tracing Malware Evolution from Embeddings at Scale

ApplicationsDGX agent

arXiv:2606.06570v1 Announce Type: cross Abstract: Malware detection remains largely reactive: machine learning models trained on known samples degrade as threats evolve. Understanding evolutionary rel

MHA-RAG: Improving Efficiency, Accuracy, and Consistency by Encoding Exemplars as Soft Prompts

ResearchDGX agent

arXiv:2510.05363v2 Announce Type: replace Abstract: Adapting Foundation Models to new domains with limited training data is challenging and computationally expensive. While prior work has demonstrated

MoDA: Modulation Adapter for Fine-Grained Visual Grounding in Instructional MLLMs

ResearchDGX agent

arXiv:2506.01850v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success in instruction-following tasks by integrating pretrained visual enco

mtmd : add video input support by ngxson · Pull Request #24269 · ggml-org/llama.cpp

Model ReleasesDGX agent

PR #24269 added native video input to llama.cpp's multimodal (mtmd) system, merging on June 8, 2026. The implementation uses FFmpeg as a subprocess to decode video frames and expands a single video ma

Multi-objective optimization and quantum hybridization of equivariant deep learning interatomic potentials

ResearchDGX agent

arXiv:2602.16908v2 Announce Type: replace-cross Abstract: Allegro is a machine learning interatomic potential model designed to predict atomic properties in molecules using E(3) equivariant neural net

OpenGlass: Open-Source Smart Glasses for On-Device Event-Based Gesture Recognition

Model ReleasesDGX agent

arXiv:2606.07431v1 Announce Type: new Abstract: Smart eyewear enables unobtrusive, context-aware interaction through multimodal sensors and on-device intelligence, but is severely limited by power, me

PaperFlow: Profiling, Recommending, and Adapting Across Daily Paper Streams

Model ReleasesDGX agent

arXiv:2606.07454v1 Announce Type: cross Abstract: Scientific paper recommendation is typically evaluated as static ranking over a fixed candidate set, yet real scientific reading unfolds as a daily, l

PhyRoGen: Synthetic Generation of Physical Robot Manipulation Puzzles Using Procedural Content Generation

Model ReleasesDGX agent

arXiv:2606.06569v1 Announce Type: new Abstract: Robot manipulation of physical puzzles is important for automatic assembly and disassembly tasks. However, to enable robots to solve physical puzzles, m

Position: Don't Just 'Fix it in Post': A Science of AI Must Study Training Dynamics

SafetyDGX agent

arXiv:2606.06533v1 Announce Type: new Abstract: What would it mean to have a scientific understanding of AI? Models are not static objects: they are snapshots of time-evolving processes shaped by data

Predictive Statistics Shape Emergent World Representations of Grid Walkers

SafetyDGX agent

arXiv:2603.16689v2 Announce Type: replace Abstract: Next-token predictors often appear to develop internal representations of the latent world and its rules. The probabilistic nature of these models s

RealDocBench: A Benchmark for Field-Level QA and Layout Understanding on Real-World Regulated Documents

Model ReleasesDGX agent

arXiv:2606.07401v1 Announce Type: new Abstract: Document parsing systems are increasingly deployed in high-stakes, regulated workflows such as mortgage underwriting, financial reporting, supply-chain

RECAP: Regression Evaluation for Continual Adaptation of Prompts

Model ReleasesDGX agent

arXiv:2606.06698v1 Announce Type: cross Abstract: Production agentic systems routinely face evolving constraints and must comply from the very next interaction. Scenarios like a tool-call notification

RISE: Single Static Radar-based Indoor Scene Understanding

Model ReleasesDGX agent

arXiv:2511.14019v3 Announce Type: replace Abstract: Robust and privacy-preserving indoor scene understanding remains a fundamental open problem. While optical sensors such as RGB and LiDAR offer high

ScatterPrism: convergence for generative simulation and inverse problems in particle and nuclear physics

ApplicationsDGX agent

arXiv:2604.01313v2 Announce Type: replace Abstract: High-fidelity simulations and complex inverse problems, such as detector modeling and unfolding, are computationally intensive bottlenecks across su

SecretFan: Synthesizing Realistic Data without Breaking Privacy

ResearchDGX agent

arXiv:2602.05833v2 Announce Type: replace Abstract: There is a need for synthetic training and test datasets that replicate statistical distributions of original datasets without compromising their co

Spatial-Temporal Decoupled Adapter for Micro-gesture Online Recognition

Model ReleasesDGX agent

arXiv:2606.07355v1 Announce Type: new Abstract: Micro-gesture online recognition aims to temporally localize and classify subtle gestures in untrimmed videos. Owing to their extremely short duration,

SV-Detect: AI-generated Text Detection with Steering Vectors

SafetyDGX agent

arXiv:2606.07313v1 Announce Type: cross Abstract: Detecting machine-generated text is especially difficult under distribution shift, such as transfer across domains, source models, and editing attacks

SWE-IF: Aligning Code Evaluation with Human Preference

ResearchDGX agent

arXiv:2510.07315v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have catalyzed vibe coding, where users leverage LLMs to generate and iteratively refine code through natural lan

The highest leverage work in AI right now is some of the most boring. (Well boring to others, I kind of love the pain of problem-solving.) E…

Model ReleasesDGX agent

The highest leverage work in AI right now is some of the most boring. (Well boring to others, I kind of love the pain of problem-solving.) Everyone and their boss wants to build the cool AI agent... t

Twin: Tuning Learning Rate and Weight Decay of Deep Homogeneous Classifiers without Validation

Model ReleasesDGX agent

arXiv:2403.05532v2 Announce Type: replace-cross Abstract: We introduce Tune without Validation (Twin), a simple and effective pipeline for tuning learning rate and weight decay of homogeneous classifi

Unsupervised Continual Clustering via Forward-Backward Knowledge Distillation

Model ReleasesDGX agent

arXiv:2606.07474v1 Announce Type: new Abstract: Unsupervised Continual Learning (UCL) aims to enable neural networks to learn sequential tasks without labels or access to past data. A major challenge

VideoSEG-O3: A Multi-turn Reinforcement Learning Framework for Reasoning Video Object Segmentation

Model ReleasesDGX agent

arXiv:2606.06819v1 Announce Type: new Abstract: Reasoning Video Object Segmentation (RVOS) demands a sophisticated integration of temporal dynamics, spatial details, and linguistic reasoning to achiev

Watch, Remember, Reason: Human-View Video Understanding with MLLMs

SafetyDGX agent

arXiv:2606.07433v1 Announce Type: cross Abstract: Video understanding is being rapidly transformed by multimodal large language models (MLLMs), as research moves from short clips to long, multimodal,

Your UnEmbedding Matrix is Secretly a Feature Lens for Text Embeddings

ResearchDGX agent

arXiv:2606.07502v1 Announce Type: new Abstract: Large language models exhibit impressive zero-shot capabilities across a wide range of downstream tasks. However, they struggle to function as off-the-s

7 Jun 2026

Grep timeout issue fixed in latest Grok Build

Model ReleasesDGX agent

Grep timeout issue fixed in latest Grok Build Grok Build update just released v0.2.31 Release Notes: Bug Fixes: • Marketplace skills without proper descriptions are now hidden from listings instead of

Zuckerberg and LeCun’s unilateral decision to open source Llama likely (partly) catalyzed China’s AI industry — and may have done truly mass…

Model ReleasesDGX agent

Zuckerberg and LeCun’s unilateral decision to open source Llama likely (partly) catalyzed China’s AI industry — and may have done truly massive harm to American business interests. We are now starting

6 Jun 2026

Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning

AgentsDGX agent

arXiv:2503.01734v3 Announce Type: replace-cross Abstract: Attacks on machine learning models have been extensively studied through stateless optimization. In this paper, we demonstrate how a reinforce

Agentic Monte Carlo: Simulating Reinforcement Learning for Black-Box Agents

Model ReleasesDGX agent

arXiv:2606.05296v1 Announce Type: cross Abstract: LLM agents operate in two distinct regimes: open-weight agents amenable to reinforcement learning (RL) and black-box agents whose behaviour must be co

Benchmarking Counterfactual Prediction in Epidemic Time Series with Time-Varying Interventions

Model ReleasesDGX agent

arXiv:2606.05692v1 Announce Type: cross Abstract: Deep learning has enabled significant advances in time-series causal inference, yet progress remains constrained by the lack of realistic benchmarks w

Brick-Composer: Using MLLMs for Assembly with Diverse Bricks

Model ReleasesDGX agent

arXiv:2606.05445v1 Announce Type: new Abstract: We dream of AI agents that can read arbitrary designs and construct real-world objects from reusable building blocks. As a first step toward this vision

CangLing-KnowFlow: A Unified Knowledge-and-Flow-fused Agent for Comprehensive Remote Sensing Applications

Model ReleasesDGX agent

arXiv:2512.15231v3 Announce Type: replace Abstract: The automated and intelligent processing of massive remote sensing (RS) datasets is critical in Earth observation (EO). Existing automated systems a

Closing the Loop on Latent Reasoning via Test-Time Reconstruction

Model ReleasesDGX agent

arXiv:2606.06252v1 Announce Type: new Abstract: Recent work moves intermediate reasoning from natural-language traces into latent or cache-level representations to reduce token overhead and avoid a di

Consistency Training Along the Transformer Stack

SafetyDGX agent

arXiv:2606.05817v1 Announce Type: cross Abstract: Consistency training encourages models to behave similarly across different contexts, and has shown promise for reducing misalignment. We broaden the

CTIConnect: A Benchmark for Retrieval-Augmented LLMs over Heterogeneous Cyber Threat Intelligence

Model ReleasesDGX agent

arXiv:2510.11974v2 Announce Type: replace-cross Abstract: Cyber Threat Intelligence (CTI) is foundational to modern cybersecurity, enabling organizations to proactively defend against evolving threats

How Far Did They Go? The Persuasive Tactics of Covert LLM Agents in a Discontinued Field Experiment

Model ReleasesDGX agent

arXiv:2606.05256v1 Announce Type: new Abstract: This study analyzes a publicly released dataset from a discontinued field experiment on Reddit's r/ChangeMyView. The intervention, conducted by unknown,

I Know What You Meme, Even If it Emerged Today: Understanding Evolving Memes through Open-World Knowledge Acquisition

Model ReleasesDGX agent

arXiv:2606.05316v1 Announce Type: new Abstract: Multimodal memes are dynamic and often require up to date background knowledge for interpretation. Existing methods often overlook such knowledge or rel

LLM Research Papers: The 2026 List (January to May)

ResearchDGX agent

This resource compiles significant large language model research papers published between January and May 2026, curated by Sebastian Raschka. It serves as a reference guide for tracking recent advance

RedKnot: Efficient Long-Context LLM Serving with Head-Aware KV Reuse and SegPagedAttention

HardwareDGX agent

arXiv:2606.06256v1 Announce Type: new Abstract: As the input length of large language model (LLM) serving continues to grow, the KV cache has become a dominant bottleneck in AI infrastructure. It limi

Self-Commitment Latency: A Reward-Free Probe for Prompted Implicit Hacking

ResearchDGX agent

arXiv:2606.05625v1 Announce Type: new Abstract: Implicit reward hacking is hard to audit when a language model's chain of thought appears benign: a final answer may be anchored by a prompt shortcut wh

← Previous
1…538539540541542…1062
Next →