How RISC-V Chips Are Becoming the Secret Weapon for Edge AI
Andes Technology's RISC-V neural processing units now run transformer and vision-language models at the edge, bringing server-grade AI to.
127 articles
Andes Technology's RISC-V neural processing units now run transformer and vision-language models at the edge, bringing server-grade AI to.
Hanwha's dual neural processing unit design delivers 3x AI inference throughput while meeting NDAA compliance, positioning it as the top Hikvision.
Perimeter Compute will install GPUs in office building basements to tap over a gigawatt of idle urban power capacity for edge AI inference.
Edge AI for IoT lets billions of devices process data locally, cutting latency to milliseconds, slashing bandwidth costs, and working offline without.
On-device AI inference is moving AI off the cloud and onto your phone, watch, and PC, delivering instant responses and stronger privacy without an.
Smartwatch neural processing units now detect heart irregularities in milliseconds, locally, as edge AI chips race toward a $129 billion market by 2033.
On-device federated learning is reshaping Silicon Valley's AI strategy, keeping data local while forcing hard trade-offs on accuracy, energy, and device.
Local AI hardware like the $4,000 DGX Spark can pay for itself in months, freeing developers from cloud bills and usage caps for on-device inference.
Verified ECAD models for BrainChip's neuromorphic AKD1500 chip are now in 25-plus design tools, cutting evaluation time from days to minutes for edge AI.
Meta's AI glasses are getting smarter, but bans on harassment videos and a facial recognition scandal reveal a privacy crisis the upgrades can't fix.
Storage speed, not processing power, is now the hidden bottleneck for on-device AI, driving new UFS 5.0 flash memory and local inference tools to close.
mimik's new AI operating system for neural processing units fits in under 20MB, letting edge devices run multiple AI agents on CPU, GPU, and NPU at once.
Qualcomm's neural processing units are reshaping edge AI, with the company targeting $40 billion in non-handset revenue by 2029 across autos, IoT, and PCs.
SAPPHIRE's EDGE+ Apex packs a neural processing unit, CPU, and GPU onto one chip, replacing the multi-board setups that have long slowed robotics.
Smart glasses sold 7 million units in 2025 while AI gadgets like the Humane Pin flopped; here's why AI wearables are finally becoming useful.
Multiverse Computing raised $570 million to compress AI models by up to 95%, enabling full on-device inference without cloud connectivity.
AMD is building an AI research hub in South Korea to link its CPUs and GPUs with Korean neural processing units in an open computing ecosystem.
Enterprise AI is moving into production, and DataGallery's open-source platform lets AI agents reason over real business data safely, ranking first among.
AI wearables can record your life without others knowing, and privacy experts warn the consent model we've relied on for decades is disintegrating.
Huawei's Kirin 9030S delivers a 200% NPU boost, enabling real-time 200MP neural processing on-device for faster, private smartphone photography.
NPUs are the new AI battleground: AMD's latest chip hits 50 TOPS locally, while startups target 100-billion-parameter models without the cloud.
On-device AI inference is projected to grow 27.84% annually through 2031, outpacing cloud, as enterprises prioritize local processing for security and.
On-device AI is splitting into three distinct approaches: medical wearables, portable robotics hubs, and mini PCs running 120-billion-parameter models.
On-device inference is helping AI companion apps cut response latency 40% and unlock enterprise adoption, in a market set to grow from $17B to $108B by.
Cisco and AMD are treating every AI PC like a secure network node, as neural processing units push enterprise AI inference from cloud servers to employee.
Battery life, not processing power, is the real limit for on-device AI in 2026; how often a model runs matters as much as its size.
AI wearables are now practical daily tools, with smart glasses, health trackers, and voice recorders solving real problems through smarter, lighter.
NVIDIA's Cosmos 3 Edge and femtoAI's SPU cut on-device inference power by up to 100X, signaling edge AI is now about precision, not just size.
Perth startup Nanoveu recorded a 27.8% drone efficiency gain using its edge AI chip, now in TSMC fabrication, but commercial orders remain unproven.
Modern browsers are running AI directly on your device, cutting latency, boosting privacy, and enabling offline use as edge AI heads toward $11.86 billion.
Samsung SDS now offers neural processing units as a subscription, letting companies run AI inference without buying servers or relying on costly GPU cloud.
NPUs promise AI acceleration in new PCs, but educators warn most users will never use the features they power, making the premium cost hard to justify.
Emdoor's Ailyn platform links PCs, NAS, and IoT devices into a unified on-device AI network that keeps data local and works offline.
On-device inference chips are set to surge from $8.4 billion to $89.6 billion by 2034, making your phone's AI brain faster, private, and battery-friendly.
AMD acquired FastFlowLM to bring fast, private on-device AI inference to PCs and workstations, cutting cloud dependency with an open-source NPU stack.
On-device AI inference is hitting factory floors, with NVIDIA compressing its 65B-parameter world model to 4B so robots can act instantly without cloud.
AMD's new Vitis AI toolkit cuts edge AI deployment from weeks to days by automating quantization and video pipeline integration for neural processing.
Edge AI devices can be hijacked by physical voltage attacks with 56% success, but a new software defense called Shuffled-ArgMax blocks them completely.
AI wearables are shifting from passive recording to active agent control, as startups like Aina raise millions to let users trigger AI workflows directly.
GDDR7 and ReRAM memory chips are turning AI inference into a trillion-dollar on-device market, with both technologies growing at 40-plus percent annually.
IBM's new Power S1112 edge server brings AI inference on-premises, cutting API costs from $8,000 to $384 annually for heavy workloads.
South Korea's Daum runs AI search on homegrown chips and models, with FuriosaAI's neural processing unit cutting costs 1.5 times versus Nvidia's H200.
Smart rings like the RingConn 3 promise AI health monitoring but fail real-world tests, missing severe migraines while reporting users are in top form.
Texas Instruments' TIDA-010997 BoosterPack enables on-device inference across motion, audio, and environmental sensors, no cloud connection required.
Quadric raised $90 million to run AI inference on-device, cutting cloud costs for SMEs with a programmable chip that updates via software as models evolve.
Samsung's GAIA neural processing unit chip could shake up the AI PC market by 2027, with prototypes already in testing at Lenovo and HP.
AI companies are racing to route tasks across smaller, cheaper models, with open-weight options potentially handling 90% of AI tokens by end of 2026.
Windows AI laptops are splitting by NPU architecture, and choosing between Intel AI Boost and AMD Ryzen AI now shapes which features run smoothly for.
Google's LiteRT.js brings on-device inference to browsers via WebGPU, running AI up to 3x faster than older web runtimes with full privacy and zero server.
Syntiant's IPO reveals that 97% of its revenue comes from sensors, not AI chips, complicating its pitch as a pure on-device inference play.