AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
6 Aug 2026

b10290

Model ReleasesDGX agent

mtmd/ggml: add ggml_build_forward_order (#26649) ggml: add ggml_build_forward_order ggml_build_forward_expand marks the tensor and all its ancestors for compute, so using it as a pure ordering hint (k

b10291

Model ReleasesDGX agent

vulkan: fix submission batching size, add debug tools for diagnosing causes of DeviceLost drivers errors (#26371) vulkan: add debug tooling to get more information about a DeviceLost error fix submiss

b10293

Model ReleasesDGX agent

ci : onboard AMD ROCm CI with gfx1151 fixes (#26544) ci: prepare for amd rocm ci Signed-off-by: Aaron Teo aaron.teo1@ibm.com ci: fix editorconfig-checker Signed-off-by: Aaron Teo aaron.teo1@ibm.com ci

b10295

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

model-loader : fix quantized reshaped tensor strides (#26672) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64)

b10297

Model ReleasesDGX agent

server: fix empty response for /cors-proxy (#26656) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFra

b10298

Model ReleasesDGX agent

mtmd: add chunk save/load function (#26645) mtmd: add chunk save/load function nits add tests rn _MAX --> _COUNT Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (

Best llama cpp flags to run Deepseek-flash 0731

Model ReleasesDGX agent

Hi all. These are my system specs: dual xeon e5 2696 v2 , 160gb DDR3 ram ECC(1600mhz), 3 gpus: 3060 12gb, p100 16gb, 3050 6gb. And a 400gb nvme sdd RAID0, 3000 mb/s. The model is Deepseek-flash-0731 U

BIM-Native Tokenization for Constraint-Aware Room Layout Synthesis

Model ReleasesDGX agent

arXiv:2512.04832v3 Announce Type: replace Abstract: We present a BIM-native tokenization for room-level layout synthesis in Building Information Modeling (BIM) scenes. The core contribution is represe

BnBERT-iPET: Sparse Few-Shot Language Modeling for Bengali via Lottery Ticket Pruning

Model ReleasesDGX agent

arXiv:2608.05104v1 Announce Type: new Abstract: Deep neural networks have shown impressive success in NLP tasks owing to their complex structure and huge number of edges. Achieving state-of-the-art pe

BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding

Model ReleasesDGX agent

arXiv:2608.04156v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instru

Breaking the Curse ofMultilinguality inMany-to-Many Speech-to-Text Translation via a Resource-AwareMixture of Speech Encoders

Model ReleasesDGX agent

arXiv:2608.04586v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved significant success in speech-to-text translation (S2TT). However, when processing multilingual

Caching for the Future: Scrub Jay Episodic Memory Principles for Agent Memory Systems

Model ReleasesDGX agent

arXiv:2608.04746v1 Announce Type: new Abstract: LLM agents that persist across sessions accumulate stored memories whose validity varies enormously by content type, yet existing memory architectures t

Can Post-Training Transform LLMs into Causal Reasoners?

Model ReleasesDGX agent

arXiv:2602.06337v2 Announce Type: replace-cross Abstract: Causal inference is essential for decision-making but remains challenging for non-experts. While large language models (LLMs) show promise in

Causal Evidence Extraction and Triangulation in Crisis Reports using Large Language Models: A ReliefWeb-based Study

Model ReleasesDGX agent

arXiv:2608.04576v1 Announce Type: new Abstract: Humanitarian reports are long, noisy, and multi-topic, making it difficult to consolidate decision-relevant causal evidence. We present a ReliefWeb stud

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings

Model ReleasesDGX agent

arXiv:2608.04735v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is increasingly treated as an important safety layer for frontier reasoning models. Most monitorability evaluations st

CIDR: A Large-Scale Industrial Source Code Dataset for Software Engineering Research

Model ReleasesDGX agent

arXiv:2605.12153v2 Announce Type: replace-cross Abstract: We present the Curated Industrial Developer Repository (CIDR), a large-scale dataset of real-world software repositories collected from indust

CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision

Model ReleasesDGX agent

arXiv:2512.22969v2 Announce Type: replace Abstract: Conventional object detectors rely on cross-entropy classification, which can be vulnerable to class imbalance and label noise. We propose CLIP-Join

CoCo-IR: Contextual Composed Image Retrieval

Model ReleasesDGX agent

arXiv:2608.05149v1 Announce Type: new Abstract: Current instruction-based image retrieval systems are powerful but limited to single-turn interactions, failing to capture the iterative nature of compl

ContextWeave: A Real-World Workflow Benchmark

Model ReleasesDGX agent

arXiv:2608.04830v1 Announce Type: new Abstract: Memory is essential as language agents move from isolated tasks to long-horizon, stateful workflows, yet existing evaluations often reduce it to retriev

Continual-Learning Physics-Informed Neural Networks for Parameterized Partial Differential Equations

Model ReleasesDGX agent

arXiv:2608.04778v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) incorporate governing equations into neural-network training and can approximate PDE solutions without requirin

Cooking beyond Frames: A Stereo Event Camera Dataset in the Kitchen

Model ReleasesDGX agent

arXiv:2608.04865v1 Announce Type: new Abstract: Event cameras, also known as neuromorphic cameras, have gained significant attention in recent years due to their high temporal resolution, high dynamic

Coupled Continuous-Discrete Generation for Scene Text Image Super-Resolution

Model ReleasesDGX agent

arXiv:2608.04525v1 Announce Type: new Abstract: Scene text image super-resolution (STISR) aims to recover visually plausible appearance while preserving character semantics from degraded inputs. Exist

DeepSeek says it plans to implement substantial price increases across its services; V4 Flash currently costs 0.14/1M input and 0.28/1M output tokens (Bloomberg)

Model ReleasesDGX agent

Bloomberg: DeepSeek says it plans to implement substantial price increases across its services; V4 Flash currently costs 0.14/1M input and 0.28/1M output tokens — DeepSeek plans to implement a signifi

DeepSeek-V4 Flash 0731 vs GPT-5.6 Luna on DeepSWE: Cost and Coding

Model ReleasesDGX agent

DeepSeek‑V4 Flash 0731 is the cheapest model on the DeepSWE board, costing about 0.10 per rollout versus GPT‑5.6 Luna’s 0.61, yet it scores a pass@1 of 53.3% compared to Luna’s 67.2%. A cascade strate

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

Model ReleasesDGX agent

arXiv:2608.05004v1 Announce Type: new Abstract: Mental health professionals have raised concerns about risks of psychological harm from interaction with large language models (LLMs), including 'delusi

Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools

Model ReleasesDGX agent

arXiv:2608.04719v1 Announce Type: new Abstract: Agent evaluations tell us that a model picked the wrong tool, but rarely why. We introduce canary tools: diagnostic probe tools planted in an agent's Mo

Differential 6-DOF Pose Estimation with Provable First-Order Immunity to Camera Calibration Errors

Model ReleasesDGX agent

arXiv:2608.04673v1 Announce Type: new Abstract: Accurate six-degree-of-freedom (6-DOF) motion estimation is essential for robotic manipulation, autonomous systems, and structural displacement monitori

Digital sovereignty in the age of AI: You don’t have to choose between control and innovation

Model ReleasesDGX agent

For enterprises and governments with strict compliance and sovereignty requirements, keeping sensitive data on-premises often means missing out on the latest AI. These organizations are managing three

Diverse and Plausible Algorithmic Recourse via Tractable Recourse Distributions

Model ReleasesDGX agent

arXiv:2608.04677v1 Announce Type: new Abstract: Algorithmic recourse seeks to help individuals reverse unfavorable automated decisions by recommending actionable changes that achieve a desired outcome

DreamWAM: Beyond RGB Future Prediction for World Action Models

Model ReleasesDGX agent

arXiv:2608.04996v1 Announce Type: new Abstract: World Action Models (WAMs) learn action-relevant representations by predicting how the observed world will evolve. Most existing WAMs define this future

Dual 3090 setup: 400 pp t/s to 1600 pp t/s on Qwen 3.6 27B... with slightly lower tps.

Model ReleasesDGX agent

First of all, my setup: Ryzen 9 5950x DDR4 3200Mhz 64gb (2x32) Dual 3090s, no NVLINK Runtime: llama.cpp Nvidia Drivers 610 Windows 11 25H2 Qwen 3.6 27B Q8 I've been using llama-server with --split-mod

Dynamic Jailbreaking Attack

Model ReleasesDGX agent

arXiv:2510.02422v4 Announce Type: replace-cross Abstract: Existing gradient-based jailbreak attacks typically optimize a fixed-length adversarial suffix toward a predefined target response with a stat

Easy to Complete, Hard to Choose: Investigating LLM Performance on the ProverbIT Benchmark

Model ReleasesDGX agent

arXiv:2608.04670v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed computational linguistics and achieved remarkable performance across numerous natural language processin

Echo Flow Networks

Model ReleasesDGX agent

arXiv:2509.24122v3 Announce Type: replace Abstract: At the heart of time-series forecasting (TSF) lies a fundamental challenge: how can models efficiently and effectively capture long-range temporal d

EDATracer: An Agentic Framework for Large-Scale EDA Artifact Analysis

Model ReleasesDGX agent

arXiv:2608.04032v1 Announce Type: cross Abstract: Modern chip design relies on electronic design automation (EDA) tools that generate large, heterogeneous artifacts, including source files, scripts, l

Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits

Model ReleasesDGX agent

arXiv:2608.04324v1 Announce Type: cross Abstract: This paper studies generalized low-rank matrix bandits with multiple prioritized objectives. At each round, the learner selects a matrix-valued arm an

EgoAfford: Task-Oriented Affordance Grounding via Egocentric Referring Segmentation

Model ReleasesDGX agent

arXiv:2608.04533v1 Announce Type: new Abstract: Part-level affordance grounding has advanced the localization of functional object regions associated with elemental actions. Extending this capability

Eliciting Intrinsic Hallucinations in LLMs via Semantically Equivalent Adversarial Attacks

Model ReleasesDGX agent

arXiv:2608.04286v1 Announce Type: new Abstract: Large language models (LLMs) are often used in conjunction with external knowledge sources to improve their factual accuracy and decrease hallucinations

Energy- and Memory-Efficient PEFT Methods for Personalized On-Device SLMs on Consumer GPUs

Model ReleasesDGX agent

arXiv:2608.04488v1 Announce Type: new Abstract: Despite rapid advances in large language models (LLMs), deploying and personalizing them on resource-constrained devices remains impractical due to high

Energy-Tweedie: Score meets Score, Energy meets Energy

Model ReleasesDGX agent

arXiv:2512.23818v2 Announce Type: replace-cross Abstract: Denoising and score estimation are classically linked through Tweedie's formula, which relates the posterior mean under Gaussian noise to the

Enforcing data residency with single-Region Claude Code on Amazon Bedrock

Model ReleasesDGX agent

A regulated customer needed all Claude Code inference processed in a single AWS Region (London), not just in-geography. This post shows two ways to pin Claude Code on Amazon Bedrock to one Region: an

Enhancing Trustworthy Clinical Diagnosis Decision-Making in Large Language Models via Etiology-Aware Attention Supervision

Model ReleasesDGX agent

arXiv:2508.00285v2 Announce Type: replace Abstract: Objective: Large Language Models (LLMs) have demonstrated strong capabilities in medical text understanding and generation. However, their trustwort

EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks

Model ReleasesDGX agent

arXiv:2608.04549v1 Announce Type: cross Abstract: Frontier LLMs are increasingly put to use on open-ended complex questions, different in nature from the ones they are typically evaluated on. We dedic

ExeCRE: Execution-Consistency Guided Reliability Estimation for Self-Correcting Code Generation

Model ReleasesDGX agent

arXiv:2608.04439v1 Announce Type: cross Abstract: Large language models (LLMs) have made notable progress in code generation, but they still struggle on challenging tasks that require sophisticated al

Faster-WAM: Efficient Inference-Time Future Conditioning for Robust World Action Models

Model ReleasesDGX agent

arXiv:2608.04404v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot manipulation by learning how the environment evolves beyond the current observation. However, existing approach

Filing: DeepSeek has invested ~$20.8M in Unitree Robotics' Shanghai IPO and agreed to jointly develop AI models for humanoid machines (Eduardo Baptista/Reuters)

Model ReleasesDGX agent

Eduardo Baptista / Reuters: Filing: DeepSeek has invested ~20.8M in Unitree Robotics' Shanghai IPO and agreed to jointly develop AI models for humanoid machines — Chinese artificial intelligence start

Final optimization: from ~10 tok/s to ~15 tok/s on DeepSeek-V4-Flash-0731 at 128K ctx - 1 RTX 3090

Model ReleasesDGX agent

J'ai consacré beaucoup de temps à l'optimisation de DeepSeek-V4-Flash-0731 GGUF sur une seule RTX 3090. Mon exigence absolue pour chaque configuration était la suivante : Le modèle doit rester utilisa

FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents

Model ReleasesDGX agent

arXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unc

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

Model ReleasesDGX agent

arXiv:2608.04077v1 Announce Type: new Abstract: Evaluating financial AI agents requires criteria aligned with real professional work. Existing rubric methods typically derive criteria from task prompt

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation

Model ReleasesDGX agent

arXiv:2608.04374v1 Announce Type: cross Abstract: Large language models can produce fluent financial analysis, but fluency alone does not establish whether a report is suitable for institutional deliv

FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation

Model ReleasesDGX agent

arXiv:2511.07322v3 Announce Type: replace-cross Abstract: While LLMs have shown great success in financial tasks like stock prediction and question answering, their application in fully automating Equ

Foreseeing the Invisible: Amodal Reconstruction of Leaf Fossil Images

Model ReleasesDGX agent

arXiv:2608.04423v1 Announce Type: new Abstract: Fossil leaves are rarely preserved whole -- sedimentary rock hides, breaks, and erodes the lamina, yet paleobotany depends on the complete shape and out

Formal Analysis and Supply Chain Security for Agentic AI Skills

Model ReleasesDGX agent

arXiv:2603.00195v2 Announce Type: replace-cross Abstract: 32 pages, 5 theorems with full proofs, 68 references, open-source tool: https://github.com/qualixar/skillfortify. v2: corrects the bibliograph

From Financial Sentiment Classification to Return Predictability: A QLoRA Benchmark of Large Language Models

Model ReleasesDGX agent

arXiv:2608.04200v1 Announce Type: cross Abstract: Financial sentiment classifiers are commonly evaluated against human labels, but strong linguistic performance does not necessarily imply economically

From Score Matrices to Football-Aware Match-State Simulation: An Auditable LLM Harness for Exact-Score Reranking

Model ReleasesDGX agent

arXiv:2608.05030v1 Announce Type: new Abstract: Football score forecasting combines a strong statistical core with a difficult contextual edge. Dynamic Poisson-family models estimate team strength, ex

From Transparent Labware Segmentation to Collision Avoidance: A Real-Time Edge-Aware Perception Pipeline

Model ReleasesDGX agent

arXiv:2608.04769v1 Announce Type: new Abstract: This paper presents an edge-aware instance segmentation framework that enables real-time robotic collision avoidance with transparent laboratory glasswa

FUSEP: A Multi-Center Benchmark for Diverse Tasks in Early Pregnancy Fetal Ultrasound Screening

Model ReleasesDGX agent

arXiv:2608.04766v1 Announce Type: cross Abstract: A large number of infants with congenital anomalies are born each year globally, especially in areas with underdeveloped medical resources. Currently,

Galaxy Phase-Space and Field-Level Cosmology: The Strength of Semi-Analytic Models

Model ReleasesDGX agent

arXiv:2512.10222v2 Announce Type: replace-cross Abstract: Semi-analytic models are a widely used approach to simulate galaxy properties within a cosmological framework, relying on simplified yet physi

GEB-Bench: Abstract Structures Told in Many Voices

Model ReleasesDGX agent

arXiv:2608.04111v1 Announce Type: cross Abstract: Can a model look at a river delta and a lightning bolt and see that they share a structure? We introduce GEB-Bench, a benchmark whose unit is an abstr

General Availability of Pinecone Nexus Proves Knowledge Drives Real Outcomes for Agentic AI

Model ReleasesDGX agent

Pinecone announced the general availability of Pinecone Nexus, a knowledge engine that converts an enterprise’s proprietary data into governed, agent‑ready knowledge delivered through a single query c

← Previous
1…2425262728…372
Next →