The Mathematics Search Engine
Mathematics News & Resources
4Mathematics is a specialist search engine for Mathematics. Discover the latest math news and mathematical content. Part of the 4SEARCH network of topic specific search engines.
Latest Articles
HFEPX | Human Feedback and Eval Paper Explorer
11+ hour, 32+ min ago (49+ words) View this page in? A focused feed for RLHF, preference data, rater protocols, agent evaluation, and LLM-as-judge research. Every paper includes structured metadata for quick triage. Join the #1 platform for finding AI training and data labeling work. Find clients and…...
Is a Face Lookup Worth Using? What You Should Know First
38+ min ago (21+ words) A face lookup is worth using when the risk, uncertainty, or time pressure is real and you have limited written details....
World Bank Group Announces WBG Pioneer Data Analysis Internship in Washington, DC - Apply by 12 August 2026
48+ min ago (401+ words) The World Bank Group has announced an opportunity for graduate students to apply for the WBG Pioneer – Data Analysis Intern position under its prestigious internship programme. The internship is based in Washington, DC, United States, and provides students with practical…...
Baseten built the fastest GLM-5.2 API on earth and the playbook tells you where inference is heading
3+ hour, 33+ min ago (621+ words) Baseten is serving Zhipu AI's GLM-5.2 at 593.7 tokens per second, roughly 12.8 times faster than the next-fastest provider. The optimization stack, NVFP4 quantization on NVIDIA Blackwell, prefill-decode disaggregation via NVIDIA Dynamo, and multi-token prediction, is a preview of how the inference compute…...
LLM-as-a-Judge
5+ hour, 53+ min ago (159+ words) Every judged response costs one extra LLM call to the judge model. Define your guardrails under the guardrails section: Criterion weights must sum to 100. overall_threshold defaults to 80 and on_failure defaults to block. A response that fails the criteria is rejected with HTTP…...
Induction Labs Photon-1 Simulates Desktops, Plays Checkers, and Models Billiard Physics From One Pretraining Run
2+ hour, 17+ min ago (362+ words) Most agents that learn from video need to know what action produced each frame. Induction Labs is arguing that this requirement is the bottleneck. Last week, they released imagination models, a foundation model architecture that pretrains on raw video with…...
The Underrated AI Tool That Lets Any LLM Watch Videos
2+ hour, 31+ min ago (930+ words) I spent this month testing the ways people actually get an LLM to "watch" a video. One of them got a 2,181-video stress test from a single user. Here is what held up. The three approaches on the table 1. Upload…...
How to Deploy VDF AI Inside an Enterprise Data Center
1+ day, 7+ hour ago (495+ words) What You Can Build A practical, layer-by-layer guide to standing up VDF AI in your own data center — the compute, networking, storage, identity, and governance layers you need, how they fit together, and what platform and infrastructure teams should plan…...
FAIRChem v2 UMA for Multidomain Atomistic Simulation across Molecules, Catalysts, Materials, Vibrations, and Molecular Dynamics
2+ hour, 53+ min ago (790+ words) In this tutorial, we explore FAIRChem v2 and the UMA universal machine-learning interatomic potential as a unified framework for atomistic simulation across molecular chemistry, catalysis, and inorganic materials. We configure an environment, authenticate with Hugging Face to access the gated UMA…...
If Authors Cannot Give A Clear Talk On Their Math Results, Their Proofs Shouldn't Be Published: Terrance Tao
3+ hour, 40+ min ago (375+ words) Even as AI-generated math proofs are beginning to proliferate across the internet, one of the most prominent mathematicians of recent times suggests we shouldn’t be rushing into publishing them. Tao, who has spent the last several months as one of…...