Top story
Introducing the Agents API
heat 95
Daily AI intelligence
Track new AI products, prediction markets, research, media, models, public equities, private funding, and the signals that matter before they become consensus.
Top story
heat 95
Market signal
0% / Polymarket
Research
arXiv
Launch
+100 momentum
Product surface
Each module is designed to become dynamic through source ingestion, scoring, human curation, and realtime updates.
Rank new AI products, agents, models, APIs, and open-source projects by real-world usefulness.
Track AI-related prediction markets, probability moves, volumes, and event risk.
Curate news, podcasts, interviews, newsletters, and video into concise signal briefs.
Summarize papers, model cards, benchmarks, evals, and safety reports for practitioners.
Map public and private AI exposure across compute, cloud, apps, infra, energy, and data.
Follow funding rounds, valuation changes, strategic investments, and ecosystem concentration.
Compare frontier, open-weight, coding, reasoning, multimodal, and specialist AI models.
Launch rankings
Product Hunt-style discovery for AI tools, models, agents, APIs, and open-source releases.
Product Hunt / 1,488 votes
AI fundraising agent that finds investors and books meetings
View launchProduct Hunt / 1,155 votes
Ship AI agents like web apps, in minutes.
View launchPrediction markets
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Polymarket
Media feed
Build and launch cloud agents with the Agents API, a managed service powered by the Codex harness for orchestration, long-running sessions, and tool use.
GPT‑Live‑1 brings natural, full-duplex voice conversations to the API, with stronger instruction following, custom voices, and telephony support.
OpenAI and GSA will offer eligible federal, state, local, and tribal governments $0 license fees, 50% off usage, and expanded cyber defense support.
Research corner
A language model normally begins training with random word embeddings: whatever 'banana' means must be learned from training corpora. I implement St. Augustine's picture of word learning, meaning by ostension, for a small masked language model (DeBERTa) trained on 10M words: before training, visually grounded tokens receive embeddings derived from the image regions they label; other tokens start random. Visual initialization leaves a measurable imprint that lasts until the end of training. At the same time, the effect remains invisible under most BabyLM benchmarks, which probe abstract grammatical knowledge: visual initialization does not affect performance there. The only zero-shot exceptio
Continuous glucose monitoring (CGM) provides high-frequency measurements of glucose dynamics and enables short-term glucose forecasting for diabetes management. Although time-series foundation models have shown strong general forecasting ability, their effectiveness for CGM prediction and the added value of multimodal dietary context remain unclear. We conduct a comprehensive empirical study using eight public CGM datasets spanning Type 1 diabetes, Type 2 diabetes, and non-diabetes populations. Under a unified protocol across multiple context lengths and prediction horizons, zero-shot foundation models did not consistently outperform strong task-specific baselines such as Elastic Net and Pat
Recursive self-improvement (RSI) enables AI systems to turn experience and feedback into persistent changes that improve both their capabilities and the process of future improvement. We first use the Headroom-Closed Index (HCI) to reveal the problems of existing LLMs, then introduce the RSI concept and its development roadmap: from improvement-execution autonomy, improvement-strategy autonomy, experience-acquisition autonomy, and environment-adaptation autonomy, to recursive meta-improvement. Next we examine RSI across scenarios (e.g., scientific discovery, embodied intelligence, software engineering), highlighting their distinct requirements and development speeds. Drawing on diverse indus
Recently, a wide range of recommendation algorithms inspired by deep learning techniques have emerged as the performance leaders on several standard recommendation benchmarks. While these algorithms were built on different DL techniques (e.g., dropouts, autoencoder), they have similar performance and even similar cost functions. This paper studies whether the models' comparable performance are sheer coincidence, or they can be unified under a single framework. We find that all linear performance leaders effectively add only a nuclear-norm based regularizer, or a Frobenius-norm based regularizer. The former ones possess a (surprising) rigid structure that limits the models' predictive power b
Many biological discovery problems require experiments to be selected sequentially under constrained budgets. CRISPR screening is a prominent example, as exhaustive perturbation testing is often infeasible and candidate perturbations must instead be prioritized over multiple experimental rounds. Despite the importance of this problem, existing benchmarks for adaptive hit discovery remain limited in scale and diversity. Here, we introduce AssayBench-Loop, a large-scale benchmark for adaptive hit discovery comprising 1,389 CRISPR screens across five phenotype categories. Beyond enabling systematic evaluation, its scale makes it possible to learn acquisition strategies across historical experim
Large language models generate fluent text that can contain unfaithful claims -- a phenomenon known as hallucination. We present a multi-signal detection pipeline combining fine-tuned DeBERTa-v3 classification, Monte Carlo (MC) Dropout uncertainty quantification, and temperature-scaled calibration for response-level hallucination detection. Evaluated on the HaluEval benchmark, our pipeline achieves F1=0.915 and AUROC=0.977 on general-domain tasks, with per-task F1 scores of 0.97 (QA), 0.96 (Summarization), and 0.82 (Dialogue). MC Dropout inference further improves accuracy to 93.2%. A context ablation study confirms the model performs genuine entailment reasoning rather than exploiting surfa
Zero-cost proxies rank architectures cheaply, but their reliability varies across search spaces. We introduce CoRA-NAS (COarse Ranking + Anchor-residual), a two-stage framework combining a static ranking prior with low-cost learning-curve refinement. CoRA-Rank aggregates capacity and structure-at-initialization proxies through an equal-weight log-rank consensus and a target-free consensus gate. CoRA-Refine samples anchors across this prior, extrapolates their early validation curves, and propagates a learned residual correction with an ExtraTrees model. The refinement uses approximately 1% of the cost of fully training the candidate set. Fully trained architecture-accuracy labels are not use
As Speech Large Language Models (speech-LLMs) become increasingly multilingual, Arabic remains significantly underrepresented, highlighting the need for dedicated infrastructure to train and evaluate Arabic speech-LLMs. To address this gap, we introduce Nuha-Speech, a comprehensive initiative to develop general-purpose Arabic speech-LLMs spanning dataset construction, model training, and systematic evaluation. Specifically, we constructed a large-scale Arabic Speech Question-Answering (SQA) corpus comprising over 1.5 million training samples to allow instruction tuning over a broad range of core speech tasks. Then, the corpus was used for supervised fine-tuning based on Qwen-Omni model varia
Solving novel view synthesis (NVS) for millimeter-wave (mmWave) radar requires a renderer that is physically faithful, complex-valued, and multi-viewpoint-tractable. No prior method achieves these three properties simultaneously. Differentiable Monte Carlo (MC) ray tracers implement the radar forward model directly with explicit material modeling and complex outputs, but do not scale to the multi-view optimization NVS demands. Optical-NVS ports of NeRF, hash grids, and 3D Gaussians train fast but discard phase and replace explicit material modeling with opaque learned features, restricting them to power-only range-azimuth (RA) magnitudes. We propose 3D Point Splatting (3DPS), the first diffe
Causal discovery aims to uncover causal structures from data and is fundamental to scientific reasoning and intervention-based decision making. Its evaluation relies heavily on structural causal models (SCMs), which specify a causal graph together with the mechanisms that generate data, yet existing studies differ substantially in graph families, mechanisms, and evaluation protocols. The emergence of causal discovery foundation models (CDFMs) further complicates evaluation: performance may reflect not only causal discovery ability, but also overlap between pretraining environments and test SCMs, making results on fixed synthetic benchmarks difficult to interpret. We introduce CausalArena, a
Spatial reasoning depends not only on metric properties such as distance, angle, and shape, but also on topological relations that remain invariant under continuous deformation. Cognitive science identifies these relations as foundational to spatial understanding, yet foundation-model evaluations largely focus on metric or viewpoint-dependent relations. We introduce MindTopo, a benchmark of topological intuition across five properties grounded in cognitive science and formal topology: continuity, separation, order, enclosure, and knots. MindTopo evaluates each property at two cognitive levels. Reasoning asks a model to identify topological relations or infer how they change. Planning instant
Automatic Music Transcription (AMT) for guitar remains limited by three challenges: existing systems often fail to capture expressive techniques such as slides, bends, and percussive hits; they often assign notes to incorrect string-fret combinations; and they are typically trained on clean recordings, limiting their generalization to noisy real-world audio. To address these challenges, we propose TART, a modular four-stage audio-to-tablature pipeline consisting of (1) an audio-to-MIDI transcription model, (2) an expressive technique classifier, (3) an audio-conditioned T5 encoder-decoder for string-fret assignment, and (4) an automated tablature generator. We evaluate TART in a zero-shot se
Artificial Intelligence does more than create a governance problem. It can also reveal where institutions have already failed to provide responsiveness, belonging, care, and accountability. Once deployed, AI becomes an intervention in those conditions. It can repair, compound, substitute for, or conceal the failures it encounters. Responsible AI must therefore evaluate both the system and the institutional rupture into which it is introduced. The move from principles to protocols is already underway. The EU AI Act, NIST AI RMF, ISO/IEC 42001, and assurance practices translate commitments into roles, requirements, records, oversight, and assessment. The harder questions are what these protoco
Agentic AI is moving from bounded task execution toward systems that retain consequential state, continue operating and adapt across task boundaries. That shift creates a control problem that current harnesses largely solve by hand: objectives, retries, verification, stopping rules and other behavioral transitions are specified externally. We propose an artificial id, an adaptive internal drive for determining whether behavior should continue, stop or change. In a minimal virtual Petri-dish experiment, a controller too small to perform general-purpose reasoning and receiving no task-specific behavioral objective develops useful control through differential persistence. The same mechanism sel
Out-of-distribution length generalization, namely to extrapolate a task from short to longer context, has been studied intensively for transformers. Here we focus on distance generalization, which probes performance when inter-token distances are changed between training and inference, while keeping a fixed context length. We construct two synthetic delay copy tasks, both involving finite distances between source and recall, where tokens are copied either fully or selectively, and test models on delays unseen during training. We address three questions: (A) Do positional encoding schemes such as RoPE and ALiBi improve distance resolution relative to no positional encoding (NoPE)? (B) How doe
Generative artificial intelligence changes how firms reach customers, but standard marketing data do not record how often users see and notice a firm's name in generated answers. We develop Generative Marketing Mix Modeling (GMMM) to estimate the causal effects of Generative Engine Optimization (GEO) and Generative Engine Marketing (GEM). For GEO, GMMM combines repeated generated answers with question counts, shares of use across generative systems, and notice probabilities. For GEM, it combines records of sponsored placements with notice probabilities. GMMM compares expected business responses under alternative treatment sequences and establishes sufficient conditions for identifying the re
Camera traps often run in the field on edge hardware with limited or no connectivity, making small, locally-deployable vision-language models (VLMs) -- not frontier-scale ones -- the practically relevant class to evaluate for species identification. We test whether models in this deployment-relevant 2--8B range carry genuine taxonomic knowledge, evaluating four such VLMs (Qwen3-VL 2B/4B/8B, Gemma3 4B) against the domain-specific specialist BioCLIP (300M parameters) on a 96-species task, comparing clean iNaturalist photographs against camera-trap imagery from 6 LILA.science collections, on two independently-sampled evaluation sets. All models identify species far above chance, but every model
As the supply of human-written text is exhausted, it has become standard practice to repeat language model training data. Prior work has studied data repetition for densely activated Transformers, but the effects of data repetition remains largely unexplored for recently dominant sparse architectures such as Mixture-of-Experts (MoE), despite their increased compute efficiency. We vary data repetition rates across single- and multi-domain data mixes, and across MoE settings, including expert count and granularity. We consistently find, for models ranging from 80M to 1B active (8.5B total) parameters, that MoEs degrade more rapidly under data repetition. This effect increases with sparsity, di
Generalization under distribution shift remains a core challenge in modern machine learning, yet existing learning bound theory is limited to narrow, idealized settings and is non-estimable from samples. In this paper, we bridge the gap between theory and practical applications. We first show that existing definition of concept shift breaks when the source and target supports mismatch. Leveraging entropic optimal transport, we propose a key notion: $γ^{*}\!$-concept shifts, and derive a general error bound unifying covariate and $γ^{*}\!$-concept shifts, which applies to broad loss functions, label spaces, and stochastic labeling. We further develop estimators for these shifts with concentra
Counterfactual regret minimization (CFR) is one of the few large numerical workloads that still runs faster on CPUs than on GPUs. Each iteration sweeps a game tree with up to billions of states in millions of small, interdependent gather and scatter steps issued through a generic tree interface. On a GPU every kernel finishes in microseconds, so kernel launches and framework dispatch dominate the run time, and prior GPU implementations have lost to optimized CPU code. We observe that for a fixed game, everything about a CFR iteration except the numerical values is known before the first iteration runs. We propose GPU-CFR, a compiler and runtime built on this observation. It compiles any game
Investment index
NVIDIA / AMD / Broadcom / TSMC / ASML / Arm
Accelerator supply, networking, packaging, and utilization remain the central AI bottleneck.
Microsoft / Amazon / Google / Oracle / CoreWeave
Distribution and committed capacity shape which model companies can scale.
OpenAI / Anthropic / Perplexity / Runway / Harvey
The application layer is splitting into consumer assistants and vertical workflow systems.
Utilities / Data centers / Cooling / Storage / Networking
Power contracts and data center approvals are turning into AI growth indicators.
Realtime plan
Start with admin-approved seed data and daily editorial briefs.
Pull RSS, arXiv, market APIs, launch feeds, and company updates every 15 minutes.
Use Supabase Realtime for odds changes, breaking items, and admin publish events.