How Retail Finance teams are using Agentic AI to protect omni-channel margins
Ask a retail CFO where the quarter’s margin is landing and you will always get a hard-won answer, born from the discipline and rigor they bring to the business. And…
Ask a retail CFO where the quarter’s margin is landing and you will always get a hard-won answer, born from the discipline and rigor they bring to the business. And…
What if autonomous coding AI agents could push your vision reasoning models above 90% accuracy with almost no manual effort? When adapting vision reasoning models to production video tasks, developers…
Your prospects leave trails across multiple sources: a founder asks “What should I use for X?” in r/SaaS while their product launches on Hacker News. Stack Overflow questions spike. A…
Posted by Wei Wei, Developer Advocate In our previous blog posts Building a board game app with TensorFlow: a new TensorFlow Lite reference app and Building a reinforcement learning agent…
Writing a high-performance GPU kernel means thinking carefully about memory (see our series on Matrix Multiplication on Blackwell). Not just what data to load, but how that data is laid…
We recently announced Genie One, the data-smart AI coworker for business users. Today, we’ll dive into how business users can access Genie from anywhere using the Genie One mobile apps…
The NVIDIA Nemotron Model Reasoning Challenge invited the Kaggle community to explore a focused question: What techniques can improve reasoning accuracy when everyone starts from the same open model, benchmark,…
Production quality assurance (QA) workflows require more than individual test execution. You must organize tests into regression suites that run as a batch, and integrate them into continuous integration and…
October 19, 2023 — Posted by Surya Kanoria, Joseph Cauteruccio, Federico Tomasi, Kamil Ciosek, Matteo Rinaldi, and Zhenwen Dai – SpotifyIntroductionMany of our music recommendation problems involve providing users with…
We gave five frontier models a hard task: rebuild the full Wan 2.1 text-to-video inference pipeline on Modular’s MAX stack – without PyTorch or diffusers – in twenty hours. This…