AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
23 Apr 2026

3D Smoke Scene Reconstruction Guided by Vision Priors from Multimodal Large Language Models

ResearchDGX agent

arXiv:2604.05687v2 Announce Type: replace Abstract: Reconstructing 3D scenes from smoke-degraded multi-view images is particularly difficult because smoke introduces strong scattering effects, view-de

A Kinematic Framework for Evaluating Pinch Configurations in Robotic Hand Design without Object or Contact Models

ResearchDGX agent

arXiv:2604.20692v1 Announce Type: new Abstract: Evaluating the pinch capability of a robotic hand is important for understanding its functional dexterity. However, many existing grasp evaluation metho

Breaking the Assistant Mold: Modeling Behavioral Variation in LLM Based Procedural Character Generation

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.03396v2 Announce Type: replace Abstract: Procedural content generation has enabled vast virtual worlds through levels, maps, and quests, but large-scale character generation remains underex

Cortex 2.0: Grounding World Models in Real-World Industrial Deployment

ApplicationsDGX agent

arXiv:2604.20246v1 Announce Type: cross Abstract: Industrial robotic manipulation demands reliable long-horizon execution across embodiments, tasks, and changing object distributions. While Vision-Lan

Early-Stage Product Line Validation Using LLMs: A Study on Semi-Formal Blueprint Analysis

Model ReleasesDGX agent

arXiv:2604.20523v1 Announce Type: cross Abstract: We study whether Large Language Models (LLMs) can perform feature model analysis operations (AOs) directly on semi-formal textual blueprints, i.e., co

Fourier Weak SINDy: Spectral Test Function Selection for Robust Model Identification

ResearchDGX agent

arXiv:2604.20141v1 Announce Type: new Abstract: We introduce Fourier Weak SINDy, a minimal noise-robust and interpretable derivative-free equation learning method that combines weak-form sparse equati

MambaLiteUNet: Cross-Gated Adaptive Feature Fusion for Robust Skin Lesion Segmentation

Model ReleasesDGX agent

arXiv:2604.20286v1 Announce Type: cross Abstract: Recent segmentation models have demonstrated promising efficiency by aggressively reducing parameter counts and computational complexity. However, the

Onboard Wind Estimation for Small UAVs Equipped with Low-Cost Sensors: An Aerodynamic Model-Integrated Filtering Approach

AgentsDGX agent

arXiv:2604.20290v1 Announce Type: new Abstract: To enable autonomous wind estimation for energy-efficient flight in small unmanned aerial vehicles (UAVs), this study proposes a method that estimates f

Pushed: DFlash implementation for llama-cpp. buun-llama-cpp/llama-server -m Qwen3.6-27B.gguf -md dflash-draft-q4_k_m.gguf --spec-type dflash

Model ReleasesDGX agent

This post demonstrates a DFlash implementation integrated with llama-cpp, showcasing a speculative decoding setup that uses Qwen 3.6-27B as the main model with a smaller draft model (dflash-draft-q4_k

Scaling Self-Play with Self-Guidance

Model ReleasesDGX agent

arXiv:2604.20209v1 Announce Type: new Abstract: LLM self-play algorithms are notable in that, in principle, nothing bounds their learning: a Conjecturer model creates problems for a Solver, and both i

SGAP-Gaze: Scene Grid Attention Based Point-of-Gaze Estimation Network for Driver Gaze

Model ReleasesDGX agent

arXiv:2604.19888v1 Announce Type: new Abstract: Driver gaze estimation is essential for understanding the driver's situational awareness of surrounding traffic. Existing gaze estimation models use dri

Supplement Generation Training for Enhancing Agentic Task Performance

Model ReleasesDGX agent

arXiv:2604.20727v1 Announce Type: cross Abstract: Training large foundation models for agentic tasks is increasingly impractical due to the high computational costs, long iteration cycles, and rapid o

22 Apr 2026

A Gesture-Based Visual Learning Model for Acoustophoretic Interactions using a Swarm of AcoustoBots

AgentsDGX agent

arXiv:2604.19643v1 Announce Type: new Abstract: AcoustoBots are mobile acoustophoretic robots capable of delivering mid-air haptics, directional audio, and acoustic levitation, but existing implementa

Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps

Model ReleasesDGX agent

arXiv:2604.19533v1 Announce Type: cross Abstract: We introduce the Cyber Defense Benchmark, a benchmark for measuring how well large language model (LLM) agents perform the core SOC analyst task of th

GeoLaux: A Benchmark for Evaluating MLLMs' Geometry Performance on Long-Step Problems Requiring Auxiliary Lines

Model ReleasesDGX agent

arXiv:2508.06226v2 Announce Type: replace Abstract: Geometry problem solving (GPS) poses significant challenges for Multimodal Large Language Models (MLLMs) in diagram comprehension, knowledge applica

llama-server -hf ggml-org/Qwen3.6-27B-GGUF --spec-default

Model ReleasesDGX agent

This post likely demonstrates running Qwen2 3.6B or 27B model in GGUF format using llama-server with default specifications, showcasing inference capabilities of quantized open-source models. The comm

Multi-modal Test-time Adaptation via Adaptive Probabilistic Gaussian Calibration

Model ReleasesDGX agent

arXiv:2604.19093v1 Announce Type: cross Abstract: Multi-modal test-time adaptation (TTA) enhances the resilience of benchmark multi-modal models against distribution shifts by leveraging the unlabeled

Personalized Benchmarking: Evaluating LLMs by Individual Preferences

SafetyDGX agent

arXiv:2604.18943v1 Announce Type: new Abstract: With the rise in capabilities of large language models (LLMs) and their deployment in real-world tasks, evaluating LLM alignment with human preferences

Welcome to the agentic BI era with Looker

Model ReleasesDGX agent

By combining the analytical depth of Looker with Google’s Agentic Data Cloud, the potential to transform how we model, interact with, and act on our data appears limitless. This week at Google Cloud N

What’s new with the Cross-Cloud Network at Next ‘26

Model ReleasesDGX agent

While generative AI sparked a revolution, the true paradigm shift is the rapid evolution from standalone AI models to multi-agent autonomous systems. In this new era, the network transcends basic conn

21 Apr 2026

A proposal for PU classification under Non-SCAR using clustering and logistic model

ResearchDGX agent

arXiv:2604.17130v1 Announce Type: cross Abstract: The present study aims to investigate a cluster cleaning algorithm that is both computationally simple and capable of solving the PU classification wh

Beyond Attack Success Rate: A Multi-Metric Evaluation of Adversarial Transferability in Medical Imaging Models

TutorialsDGX agent

arXiv:2604.16532v1 Announce Type: new Abstract: While deep learning systems are becoming increasingly prevalent in medical image analysis, their vulnerabilities to adversarial perturbations raise seri

Beyond Binary Contrast: Modeling Continuous Skeleton Action Spaces with Transitional Anchors

ResearchDGX agent

arXiv:2604.17914v1 Announce Type: new Abstract: Self-supervised contrastive learning has emerged as a powerful paradigm for skeleton-based action recognition by enforcing consistency in the embedding

Chronax: A Jax Library for Univariate Statistical Forecasting and Conformal Inference

ApplicationsDGX agent

arXiv:2604.16719v1 Announce Type: new Abstract: Time-series forecasting is central to many scientific and industrial domains, such as energy systems, climate modeling, finance, and retail. While forec

Common Corpus: The Largest Collection of Ethical Data for LLM Pre-Training

ApplicationsDGX agent

arXiv:2506.01732v2 Announce Type: replace Abstract: Large Language Models (LLMs) are pre-trained on large data from different sources and domains. These datasets often contain trillions of tokens, inc

Contrastive Analysis of Linguistic Representations in Large Language Model Outputs through Structured Synthetic Data Generation and Abstracted N-gram Associations

SafetyDGX agent

arXiv:2604.17398v1 Announce Type: new Abstract: We present a methodological framework to discover linguistic and discursive patterns associated to different social groups through contrastive synthetic

Do LLMs Encode Functional Importance of Reasoning Tokens?

ResearchDGX agent

arXiv:2601.03066v2 Announce Type: replace Abstract: Large language models solve complex tasks by generating long reasoning chains, achieving higher accuracy at the cost of increased computational cost

Freshness-Aware Prioritized Experience Replay for LLM/VLM Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.16918v1 Announce Type: new Abstract: Reinforcement Learning (RL) has achieved impressive success in post-training Large Language Models (LLMs) and Vision-Language Models (VLMs), with on-pol

L1 Regularization Paths in Linear Models by Parametric Gaussian Message Passing

ResearchDGX agent

arXiv:2604.16949v1 Announce Type: new Abstract: The paper considers the computation of L1 regularization paths in a state space setting, which includes L1 regularized Kalman smoothing, linear SVM, LAS

Leveraging Kernel Symmetry for Joint Compression and Error Mitigation in Edge Model Transfer

ResearchDGX agent

arXiv:2604.17371v1 Announce Type: cross Abstract: This paper investigates communication-efficient neural network transmission by exploiting structured symmetry constraints in convolutional kernels. In

Local Inconsistency Resolution: The Interplay between Attention and Control in Probabilistic Models

ResearchDGX agent

arXiv:2604.17140v1 Announce Type: cross Abstract: We present a generic algorithm for learning and approximate inference with an intuitive epistemic interpretation: iteratively focus on a subset of the

Modeling Biomechanical Constraint Violations for Language-Agnostic Lip-Sync Deepfake Detection

ResearchDGX agent

arXiv:2604.16808v1 Announce Type: new Abstract: Current lip-sync deepfake detectors rely on pixel-level artifacts or audio-visual correspondence, failing to generalize across languages because these c

Modeling User Exploration Saturation: When Recommender Systems Should Stop Pushing Novelty

SafetyDGX agent

arXiv:2604.16419v1 Announce Type: cross Abstract: Fairness-aware recommender systems often mitigate bias by increasing exposure to under-represented or long-tail content, commonly through mechanisms t

MoRI: Learning Motivation-Grounded Reasoning for Scientific Ideation in Large Language Models

AgentsDGX agent

arXiv:2603.19044v2 Announce Type: replace Abstract: Scientific ideation aims to propose novel solutions within a given scientific context. Existing LLM-based agentic approaches emulate human research

OmniHuman: A Large-scale Dataset and Benchmark for Human-Centric Video Generation

Model ReleasesDGX agent

arXiv:2604.18326v1 Announce Type: new Abstract: Recent advancements in audio-video joint generation models have demonstrated impressive capabilities in content creation. However, generating high-fidel

Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots

AgentsDGX agent

arXiv:2505.19237v2 Announce Type: replace-cross Abstract: Self-recognition -- the ability to maintain an internal representation of one's own body within the environment -- underpins intelligent, auto

Speculative Verification: Exploiting Information Gain to Refine Speculative Decoding

SafetyDGX agent

arXiv:2509.24328v2 Announce Type: replace Abstract: LLMs have low GPU efficiency and high latency due to autoregressive decoding. Speculative decoding (SD) mitigates this using a small draft model to

TagaVLM: Topology-Aware Global Action Reasoning for Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2603.02972v2 Announce Type: replace Abstract: Vision-Language Navigation (VLN) presents a unique challenge for Large Vision-Language Models (VLMs) due to their inherent architectural mismatch: V

Tailoring Diagnostic Modeling to Individual Learners: Personalized Distractor Generation via MCTS-Guided Reasoning Reconstruction

ResearchDGX agent

arXiv:2508.11184v2 Announce Type: replace Abstract: Distractors-incorrect yet plausible answer choices in multiple-choice questions (MCQs)-are vital in educational assessments, as they help identify s

TeleEmbedBench: A Multi-Corpus Embedding Benchmark for RAG in Telecommunications

Model ReleasesDGX agent

arXiv:2604.17778v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in the telecommunications domain for critical tasks, relying heavily on Retrieval-Augmented Gener

Towards a Data-Parameter Correspondence for LLMs: A Preliminary Discussion

Model ReleasesDGX agent

arXiv:2604.17384v1 Announce Type: new Abstract: Large language model optimization has historically bifurcated into isolated data-centric and model-centric paradigms: the former manipulates involved sa

Video Panels for Long Video Understanding

Model ReleasesDGX agent

arXiv:2509.23724v2 Announce Type: replace Abstract: Recent Video-Language Models (VLMs) achieve promising results on long-video understanding, but their performance still lags behind that achieved on

When Helpers Become Hazards: A Benchmark for Analyzing Multimodal LLM-Powered Safety in Daily Life

Model ReleasesDGX agent

arXiv:2601.04043v2 Announce Type: replace Abstract: As Multimodal Large Language Models (MLLMs) become an indispensable assistant in human life, the unsafe content generated by MLLMs poses a danger to

20 Apr 2026

AgentV-RL: Scaling Reward Modeling with Agentic Verifier

AgentsDGX agent

arXiv:2604.16004v1 Announce Type: cross Abstract: Verifiers have been demonstrated to enhance LLM reasoning via test-time scaling (TTS). Yet, they face significant challenges in complex domains. Error

Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation

Model ReleasesDGX agent

arXiv:2604.15388v1 Announce Type: cross Abstract: Recent advances in large language models have improved code generation, but their use in hardware description languages is still limited. Moreover, tr

FS-Researcher: Test-Time Scaling for Long-Horizon Research Tasks with File-System-Based Agents

Model ReleasesDGX agent

arXiv:2602.01566v2 Announce Type: replace Abstract: Deep research is emerging as a representative long-horizon task for large language model (LLM) agents. However, long trajectories in deep research o

LACE: Lattice Attention for Cross-thread Exploration

ResearchDGX agent

arXiv:2604.15529v1 Announce Type: new Abstract: Current large language models reason in isolation. Although it is common to sample multiple reasoning paths in parallel, these trajectories do not inter

OjaKV: Context-Aware Online Low-Rank KV Cache Compression

Model ReleasesDGX agent

arXiv:2509.21623v2 Announce Type: replace-cross Abstract: The expanding long-context capabilities of large language models are constrained by a significant memory bottleneck: the key-value (KV) cache

Ollama we love you 👨‍💻🚀

Model ReleasesDGX agent

Ollama we love you 👨‍💻🚀 Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k2.6:cloud Try it with Hermes Agent: o

OSCBench: Benchmarking Object State Change in Text-to-Video Generation

Model ReleasesDGX agent

arXiv:2603.11698v2 Announce Type: replace-cross Abstract: Text-to-video (T2V) generation models have made rapid progress in producing visually high-quality and temporally coherent videos. However, exi

Persona-Assigned Large Language Models Exhibit Human-Like Motivated Reasoning

SafetyDGX agent

arXiv:2506.20020v2 Announce Type: replace Abstract: Reasoning in humans is prone to biases due to underlying motivations like identity protection, that undermine rational decision-making and judgment.

Privacy-Preserving LLMs Routing

ResearchDGX agent

arXiv:2604.15728v1 Announce Type: cross Abstract: Large language model (LLM) routing has emerged as a critical strategy to balance model performance and cost-efficiency by dynamically selecting servic

Sample Complexity Bounds for Stochastic Shortest Path with a Generative Model

SafetyDGX agent

arXiv:2604.16111v1 Announce Type: new Abstract: We study the sample complexity of learning an epsilon-optimal policy in the Stochastic Shortest Path (SSP) problem. We first derive sample complexity bo

Scaling Behaviors of LLM Reinforcement Learning Post-Training: An Empirical Study in Mathematical Reasoning

ResearchDGX agent

arXiv:2509.25300v4 Announce Type: replace-cross Abstract: While scaling laws for large language models (LLMs) during pre-training have been extensively studied, their behavior under reinforcement lear

Synthetic data in cryptocurrencies using generative models

ResearchDGX agent

arXiv:2604.16182v1 Announce Type: cross Abstract: Data plays a fundamental role in consolidating markets, services, and products in the digital financial ecosystem. However, the use of real data, espe

Taming Asynchronous CPU-GPU Coupling for Frequency-aware Latency Estimation on Mobile Edge

HardwareDGX agent

arXiv:2604.15357v1 Announce Type: cross Abstract: Precise estimation of model inference latency is crucial for time-critical mobile edge applications, enabling devices to calculate latency margins aga

TriagerX: Dual Transformers for Bug Triaging Tasks with Content and Interaction Based Rankings

ApplicationsDGX agent

arXiv:2508.16860v2 Announce Type: replace-cross Abstract: Pretrained Language Models or PLMs are transformer-based architectures that can be used in bug triaging tasks. PLMs can better capture token s

18 Apr 2026

easyaligner: Forced alignment with GPU acceleration and flexible text normalization (compatible with all w2v2 models on HF Hub) [P]

SafetyDGX agent

easyaligner is a forced alignment library designed to be performant and easy to use , leveraging GPU acceleration to align audio with text transcriptions. The tool supports flexible text normalization

Paper from Kimi: Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter [R]

ResearchDGX agent

Mooncake is the serving platform for Kimi developed by Moonshot AI, featuring a KVCache-centric disaggregated architecture that separates prefill and decoding clusters while leveraging underutilized C

17 Apr 2026

A Robust Approach for LiDAR-Inertial Odometry Without Sensor-Specific Modeling

ResearchDGX agent

arXiv:2509.06593v2 Announce Type: replace Abstract: Accurate odometry is a critical component in a robotic navigation stack, and subsequent modules such as planning and control often rely on an estima

← Previous
1…240241242243244…1018
Next →