AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
15 Apr 2026

From Attenuation to Attention: Variational Information Flow Manipulation for Fine-Grained Visual Perception

ResearchDGX agent

arXiv:2604.12508v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated impressive capabilities in general visual understanding, they frequently falter in fine

Frontier-Eng: Benchmarking Self-Evolving Agents on Real-World Engineering Tasks with Generative Optimization

Model ReleasesDGX agent

arXiv:2604.12290v1 Announce Type: new Abstract: Current LLM agent benchmarks, which predominantly focus on binary pass/fail tasks such as code generation or search-based question answering, often negl

GRADE: Probing Knowledge Gaps in LLMs through Gradient Subspace Dynamics

ApplicationsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.02830v2 Announce Type: replace Abstract: Detecting whether a model's internal knowledge is sufficient to correctly answer a given question is a fundamental challenge in deploying responsibl

Hermes is so fucking good, everything is smooth, work grreat, fast, learning and wow, cheated

ResearchDGX agent

Nous Research's Hermes is a series of fine-tuned large language models known for strong instruction-following, reasoning, and conversational capabilities. The post appears to be enthusiastic user feed

IMSE: Intrinsic Mixture of Spectral Experts Fine-tuning for Test-Time Adaptation

Model ReleasesDGX agent

arXiv:2603.07926v3 Announce Type: replace-cross Abstract: Test-time adaptation (TTA) has been widely explored to prevent performance degradation when test data differ from the training distribution. H

Incentivizing High-Quality Human Annotations with Golden Questions

SafetyDGX agent

arXiv:2505.19134v2 Announce Type: replace-cross Abstract: Human-annotated data plays a vital role in training large language models (LLMs), such as supervised fine-tuning and human preference alignmen

Influence Strength Estimation in Hyperbolic Space for Social Influence Maximization

ApplicationsDGX agent

arXiv:2502.13571v2 Announce Type: replace-cross Abstract: The Influence Maximization (IM) problem aims to find a small set of influential users to maximize their influence spread in a social network.

Is Fooocus Still The Best? Yup!

Local AiDGX agent

This r/StableDiffusion post appears to be a community discussion reaffirming Fooocus as a top-tier Stable Diffusion WebUI, particularly praised for its beginner-friendliness and output quality. Fooocu

LLMs Struggle with Abstract Meaning Comprehension More Than Expected

ResearchDGX agent

arXiv:2604.12018v1 Announce Type: cross Abstract: Understanding abstract meanings is crucial for advanced language comprehension. Despite extensive research, abstract words remain challenging due to t

LTX 2.3 Outpaint LoRA Test

Local AiDGX agent

This Reddit post from r/StableDiffusion showcases community testing of an outpainting LoRA built on top of the LTX-2.3 video generation model. The LoRA is designed to extend the canvas of an input vid

MISID: A Multimodal Multi-turn Dataset for Complex Intent Recognition in Strategic Deception Games

Model ReleasesDGX agent

arXiv:2604.12700v1 Announce Type: new Abstract: Understanding human intent in complex multi-turn interactions remains a fundamental challenge in human-computer interaction and behavioral analysis. Whi

Narrative-Driven Paper-to-Slide Generation via ArcDeck

Model ReleasesDGX agent

arXiv:2604.11969v1 Announce Type: new Abstract: We introduce ArcDeck, a multi-agent framework that formulates paper-to-slide generation as a structured narrative reconstruction task. Unlike existing m

Ollama cloud + GLM 5.1 slow and stupid or am I?

Local AiDGX agent

This Reddit thread likely discusses user frustrations with performance issues when running GLM-5.1 via Ollama's cloud inference option (`glm-5.1:cloud`), a flagship agentic coding model from Z.AI. Com

On the continuum limit of t-SNE for data visualization

Model ReleasesDGX agent

arXiv:2604.12041v1 Announce Type: cross Abstract: This work is concerned with the continuum limit of a graph-based data visualization technique called the t-Distributed Stochastic Neighbor Embedding (

Please, god can someone point me to a good source for creating modelfiles for specific archs?

Local AiDGX agent

This r/ollama Reddit post reflects a common community frustration around finding clear, architecture-specific documentation for writing Ollama Modelfiles. A Modelfile is the blueprint used to create a

Policy-Invisible Violations in LLM-Based Agents

Model ReleasesDGX agent

arXiv:2604.12177v1 Announce Type: new Abstract: LLM-based agents can execute actions that are syntactically valid, user-sanctioned, and semantically appropriate, yet still violate organizational polic

Qwen3 technical arch

Local AiDGX agent

Qwen3 is a series of large language models spanning both dense and Mixture-of-Experts (MoE) architectures, with parameter scales ranging from 0.6B to 235B, and a key innovation being the integration o

Refined Differentially Private Linear Regression via Extension of a Free Lunch Result

ResearchDGX agent

arXiv:2604.11820v1 Announce Type: cross Abstract: As data-privacy regulations tighten and statistical models are increasingly deployed on sensitive human-sourced data, privacy-preserving linear regres

RPG-SAM: Reliability-Weighted Prototypes and Geometric Adaptive Threshold Selection for Training-Free One-Shot Polyp Segmentation

Model ReleasesDGX agent

arXiv:2603.07436v2 Announce Type: replace Abstract: Training-free one-shot segmentation offers a scalable alternative to expert annotations where knowledge is often transferred from support images and

Socrates Loss: Unifying Confidence Calibration and Classification by Leveraging the Unknown

Model ReleasesDGX agent

arXiv:2604.12245v1 Announce Type: cross Abstract: Deep neural networks, despite their high accuracy, often exhibit poor confidence calibration, limiting their reliability in high-stakes applications.

StoryScope: Investigating idiosyncrasies in AI fiction

Model ReleasesDGX agent

arXiv:2604.03136v4 Announce Type: replace Abstract: As AI-generated fiction becomes increasingly prevalent, questions of authorship and originality are becoming central to how written work is evaluate

Subspace-Guided Feature Reconstruction for Unsupervised Anomaly Localization

Model ReleasesDGX agent

arXiv:2309.13904v3 Announce Type: replace Abstract: Unsupervised anomaly localization aims to identify anomalous regions that deviate from normal sample patterns. Most recent methods perform feature m

TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs

SafetyDGX agent

arXiv:2604.12232v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed across diverse domains, yet their vulnerability to jailbreak attacks, where adversarial inputs

The ‘Goldilocks zone’: How the AI factory ends the cycle of rebuilding pipelines from scratch

IndustryDGX agent

The AI bottleneck isn’t the model — it’s everything that has to happen to the data before the model can even touch it. Conversational analytics is now emerging as the bridge to turn already-curated da

Thermodynamic Liquid Manifold Networks: Physics-Bounded Deep Learning for Solar Forecasting in Autonomous Off-Grid Microgrids

AgentsDGX agent

arXiv:2604.11909v1 Announce Type: cross Abstract: The stable operation of autonomous off-grid photovoltaic systems requires solar forecasting algorithms that respect atmospheric thermodynamics. Contem

Token-Level Policy Optimization: Linking Group-Level Rewards to Token-Level Aggregation via Sequence-Level Likelihood

SafetyDGX agent

arXiv:2604.12736v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has significantly advanced the reasoning ability of large language models (LLMs), particularly in their mathem

Towards Long-horizon Agentic Multimodal Search

Model ReleasesDGX agent

arXiv:2604.12890v1 Announce Type: cross Abstract: Multimodal deep search agents have shown great potential in solving complex tasks by iteratively collecting textual and visual evidence. However, mana

Towards Realistic and Consistent Orbital Video Generation via 3D Foundation Priors

ResearchDGX agent

arXiv:2604.12309v1 Announce Type: new Abstract: We present a novel method for generating geometrically realistic and consistent orbital videos from a single image of an object. Existing video generati

Tree Learning: A Multi-Skill Continual Learning Framework for Humanoid Robots

Model ReleasesDGX agent

arXiv:2604.12909v1 Announce Type: new Abstract: As reinforcement learning for humanoid robots evolves from single-task to multi-skill paradigms, efficiently expanding new skills while avoiding catastr

Visual Preference Optimization with Rubric Rewards

Model ReleasesDGX agent

arXiv:2604.13029v1 Announce Type: cross Abstract: The effectiveness of Direct Preference Optimization (DPO) depends on preference data that reflect the quality differences that matter in multimodal ta

VPTracker: Global Vision-Language Tracking via Visual Prompt

Local AiDGX agent

arXiv:2512.22799v2 Announce Type: replace Abstract: Vision-Language Tracking aims to continuously localize objects described by a visual template and a language description. Existing methods, however,

X-VC: Zero-shot Streaming Voice Conversion in Codec Space

Model ReleasesDGX agent

arXiv:2604.12456v1 Announce Type: cross Abstract: Zero-shot voice conversion (VC) aims to convert a source utterance into the voice of an unseen target speaker while preserving its linguistic content.

14 Apr 2026

ACE-Bench: A Lightweight Benchmark for Evaluating Azure SDK Usage Correctness

Model ReleasesDGX agent

arXiv:2604.09564v1 Announce Type: cross Abstract: We present ACE-Bench (Azure SDK Coding Evaluation Benchmark), an execution-free benchmark that provides fast, reproducible pass or fail signals for wh

Adversarial Video Promotion Against Text-to-Video Retrieval

ResearchDGX agent

arXiv:2508.06964v3 Announce Type: replace Abstract: Thanks to the development of cross-modal models, text-to-video retrieval (T2VR) is advancing rapidly, but its robustness remains largely unexamined.

Agent^2 RL-Bench: Can LLM Agents Engineer Agentic RL Post-Training?

Model ReleasesDGX agent

arXiv:2604.10547v1 Announce Type: new Abstract: We introduce Agent^2 RL-Bench, a benchmark for evaluating agentic RL post-training -- whether LLM agents can autonomously design, implement, and run com

Agents in Ollama and Langflow

Local AiDGX agent

This Reddit post from r/ollama likely discusses how to build and run AI agents locally by combining Ollama — which handles local model serving to keep data private — with Langflow's visual, drag-and-d

Are Pretrained Image Matchers Good Enough for SAR-Optical Satellite Registration?

ResearchDGX agent

arXiv:2604.10217v1 Announce Type: new Abstract: Cross-modal optical-SAR (Synthetic Aperture Radar) registration is a bottleneck for disaster-response via remote sensing, yet modern image matchers are

At FullTilt: Real-Time Open-Set 3D Macromolecule Detection Directly from Tilted 2D Projections

Local AiDGX agent

arXiv:2604.10766v1 Announce Type: new Abstract: Open-set 3D macromolecule detection in cryogenic electron tomography eliminates the need for target-specific model retraining. However, strict VRAM cons

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels

SafetyDGX agent

arXiv:2604.10367v1 Announce Type: new Abstract: Audio-driven human video generation has achieved remarkable success in monologue scenarios, largely driven by advancements in powerful video generation

Beyond the Beep: Scalable Collision Anticipation and Real-Time Explainability with BADAS-2.0

Model ReleasesDGX agent

arXiv:2604.05767v2 Announce Type: replace-cross Abstract: We present BADAS-2.0, the second generation of our collision anticipation system, building on BADAS-1.0, which showed that fine-tuning V-JEPA2

BiT-MCTS: A Theme-based Bidirectional MCTS Approach to Chinese Fiction Generation

ResearchDGX agent

arXiv:2603.14410v3 Announce Type: replace Abstract: Generating long-form linear fiction from open-ended themes remains a major challenge for large language models, which frequently fail to guarantee g

BITS Pilani at SemEval-2026 Task 9: Structured Supervised Fine-Tuning with DPO Refinement for Polarization Detection

Model ReleasesDGX agent

arXiv:2604.11121v1 Announce Type: new Abstract: The POLAR SemEval-2026 Shared Task aims to detect online polarization and focuses on the classification and identification of multilingual, multicultura

BlasBench: An Open Benchmark for Irish Speech Recognition

Model ReleasesDGX agent

arXiv:2604.10736v1 Announce Type: new Abstract: No open Irish-specific benchmark compares end-user ASR systems under a shared Irish-aware evaluation protocol. To solve this, we release BlasBench, an o

Both Ends Count! Just How Good are LLM Agents at 'Text-to-Big SQL'?

Model ReleasesDGX agent

arXiv:2602.21480v4 Announce Type: replace-cross Abstract: Text-to-SQL and Big Data are both extensively benchmarked fields, yet there is limited research that evaluates them jointly. In the real world

Bottleneck Tokens for Unified Multimodal Retrieval

ResearchDGX agent

arXiv:2604.11095v1 Announce Type: cross Abstract: Adapting decoder-only multimodal large language models (MLLMs) for unified multimodal retrieval faces two structural gaps. First, existing methods rel

Catalog-Native LLM: Speaking Item-ID Dialect with Less Entanglement for Recommendation

ResearchDGX agent

arXiv:2510.05125v2 Announce Type: replace Abstract: While collaborative filtering delivers predictive accuracy and efficiency, and Large Language Models (LLMs) enable expressive and generalizable reas

CocoaBench: Evaluating Unified Digital Agents in the Wild

Model ReleasesDGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

CoFusion: Multispectral and Hyperspectral Image Fusion via Spectral Coordinate Attention

Model ReleasesDGX agent

arXiv:2604.10584v1 Announce Type: new Abstract: Multispectral and Hyperspectral Image Fusion (MHIF) aims to reconstruct high-resolution images by integrating low-resolution hyperspectral images (LRHSI

COMPOSITE-Stem

Model ReleasesDGX agent

arXiv:2604.09836v1 Announce Type: new Abstract: AI agents hold growing promise for accelerating scientific discovery; yet, a lack of frontier evaluations hinders adoption into real workflows. Expert-w

Context-Aware Semantic Segmentation via Stage-Wise Attention

Model ReleasesDGX agent

arXiv:2601.11310v2 Announce Type: replace Abstract: Semantic ultra-high-resolution (UHR) image segmentation is essential in remote sensing applications such as aerial mapping and environmental monitor

Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training

AgentsDGX agent

arXiv:2507.15640v2 Announce Type: replace-cross Abstract: Continual pre-training on small-scale task-specific data is an effective method for improving large language models in new target fields, yet

DDO-RM for LLM Preference Optimization: A Minimal Held-Out Benchmark against DPO

Model ReleasesDGX agent

arXiv:2604.11119v1 Announce Type: cross Abstract: This paper reorganizes the current manuscript around the DPO versus DDO-RM preference-optimization project and focuses on two parts: the algorithmic v

DeepSketcher: Internalizing Visual Manipulation for Multimodal Reasoning

ResearchDGX agent

arXiv:2509.25866v2 Announce Type: replace Abstract: The 'thinking with images' paradigm represents a pivotal shift in the reasoning of Vision Language Models (VLMs), moving from text-dominant chain-of

Differentially Private Verification of Distribution Properties

Model ReleasesDGX agent

arXiv:2604.10819v1 Announce Type: cross Abstract: A recent line of work initiated by Chiesa and Gur and further developed by Herman and Rothblum investigates the sample and communication complexity of

Diffusion-Based Generative Priors for Efficient Beam Alignment in Directional Networks

SafetyDGX agent

arXiv:2604.09653v1 Announce Type: cross Abstract: Beam alignment is a key challenge in directional mmWave and THz systems, where narrow beams require accurate yet low-overhead training. Existing learn

Diffusion-CAM: Faithful Visual Explanations for dMLLMs

ResearchDGX agent

arXiv:2604.11005v1 Announce Type: new Abstract: While diffusion Multimodal Large Language Models (dMLLMs) have recently achieved remarkable strides in multimodal generation, the development of interpr

Discourse Diversity in Multi-Turn Empathic Dialogue

ResearchDGX agent

arXiv:2604.11742v1 Announce Type: cross Abstract: Large language models (LLMs) produce responses rated as highly empathic in single-turn settings (Ayers et al., 2023; Lee et al., 2024), yet they are a

Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations

SafetyDGX agent

arXiv:2604.11322v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive capabilities in utilizing external tools. In practice, however, LLMs are often exposed to to

EDUMATH: Generating Standards-aligned Educational Math Word Problems

ApplicationsDGX agent

arXiv:2510.06965v2 Announce Type: replace-cross Abstract: Math word problems (MWPs) are critical K-12 educational tools, and customizing them to students' interests and ability levels can enhance lear

Endogenous Information in Routing Games: Memory-Constrained Equilibria, Recall Braess Paradoxes, and Memory Design

SafetyDGX agent

arXiv:2604.11733v1 Announce Type: cross Abstract: We study routing games in which travelers optimize over routes that are remembered or surfaced, rather than over a fixed exogenous action set. The pap

← Previous
1…496497498499500…1071
Next →