CUDA & Compute

CUDA, creator workloads, AI frameworks, and encoder/decoder blocks.

  • 17 Tracked terms
  • Last 30 days Feed window

What this topic collects on

An article joins this feed when it matches these terms. Each one is also a search of its own.

Latest in CUDA & Compute

DEV Community
dev.to > gde > gemma-4-on-an-old-4-gb-laptop-gpu-qat-takes-it-from-95-gib-to-16-b5l

Gemma 4 on an Old 4 GB Laptop GPU: QAT Takes It From 9.5 GiB to 1.6

2+ day, 10+ hour ago   (1449+ words) This article provides a step by step deployment guide for Gemma 4 E2B's quantization-aware-trained (QAT) checkpoint to a local, laptop hosted GPU enabled system — a much older Lenovo Yoga 9 with a 4 GB GTX 1650 Ti. A suite of Python MCP tools is…...

ETMM
etmm-online.com-online.com

Solidcam 2026 Boosts CAM Simulation With Machine Works GPU

5+ day, 19+ hour ago   (146+ words) GPU-based simulation allows users to review large machining jobs more quickly. Solidcam 2026 integrates Machine Works GPU to generate in-process stock models in seconds, helping programmers check selected stages before running a full simulation. For the majority of machining operations Machine…...

@daytonaio
daytona.io > changelog > mi355x-gpu-type-and-python-sdk-s3-upload-reliability

MI355X GPU type and Python SDK S3 upload reliability

1+ week, 2+ day ago   (48+ words) Daytona 0.210.0 adds a new GPU type to the API client and hardens S3 uploads in the Python SDK. API client: add MI355X GPU type Python SDK: improve S3 upload reliability sync go.sum for v0.207.1 Go SDK: bump to v0.210.0 New Partner with us © 2026 Daytona…...

@phoronix
phoronix.com > news > Fedora-47-Considers-Thin-LTO

Fedora 47 Considering Use Of Thin LTO Compiler Optimizations

1+ week, 1+ day ago   (227+ words) A change proposal filed for what would be part of Fedora Linux 47 is on making use of Thin LTO rather than Fat LTO for link-time optimizations. With Thin LTO being more memory efficient and faster build speeds, the hope is…...

@phoronix
phoronix.com > news > GCC-17-RISC-V-mtune-mcpu-native

GCC 17 Now Supports Using -mtune=native -mcpu=native On RISC-V

1+ week, 2+ day ago   (214+ words) As a follow up to last month's article about patches being posted for enabling "-mcpu=native -mtune=native" support for RISC-V with the GCC compiler, that code is now merged for what will become the GCC 17.1 release in the early…...

OpenTrain AI
opentrain.ai > jobs > cuda-to-python-code-translation-expert--cmsg0j6gc003b04jo4hma3q0v

CUDA to Python Code Translation Expert

1+ week, 3+ day ago   (539+ words) View this page in? Use your CUDA, C++, PyTorch, and NumPy expertise to translate GPU code, evaluate LLM-generated implementations, and improve AI through RLHF. This contractor role requires 20+ hours weekly and fluent English. Open to applicants in Create a free…...

ServeTheHome
servethehome.com > nvidia-risc-v-for-nvidia-gpus-slide-09

NVIDIA CUDA applications rely on data transfer between memories

2+ week, 6+ day ago   (19+ words) ServeTheHome NVIDIA CUDA applications rely on data transfer between memories...

ServeTheHome
servethehome.com > nvidia-risc-v-for-nvidia-gpus-slide-04

NVIDIA CUDA applications combine CPU and GPU SW modules

2+ week, 6+ day ago   (19+ words) ServeTheHome NVIDIA CUDA applications combine CPU and GPU SW modules...

krepsiniozinios
lovehouzzaq.click > product_tag > 17081590_.html

Linux Install Pytorch On Ubuntu Deep Learning Installing Pytorch In Ubuntu Installing PyTorch And TensorFlow On The Linux-based Operating System: A Complete Setup Guide

1+ week, 4+ day ago   (38+ words) krepsiniozinios Your shopping cart is empty! SAVE UP TO 50% FREE SHIPPING OVER $50...

StorageReview.com
storagereview.com > news > aws-launches-graviton5-based-ec2-r9g-and-r9gd-memory-optimized-instances

AWS Launches Graviton5-Based EC2 R9g and R9gd Memory-Optimized Instances

1+ week, 4+ day ago   (377+ words) AWS has made its Amazon EC2 R9g and R9gd instances generally available, introducing memory-optimized systems powered by AWS Graviton5 processors. The new Arm-based instances target database, caching, analytics, container, and microservices workloads that require high memory bandwidth and improved per-vCPU performance. The R9gd family provides…...