Finally, bare-metal NPU programming! Immediately thought of those custom, weird inference tasks Intel won't ever officially support.
Finally, bare-metal NPU programming! Immediately thought of those custom, weird inference tasks Intel won't ever officially support.
Finally npus can be like cuda. Why did Intel limit themselves to openvino? Didn't want to take on the burden of something like cuda? Maybe they want to separate consumer and business NPUs from data center products which use oneapi which is more like cuda
Seems that Intel prefer to provide a unified interface across different hardwares so hiding the low-level stuff is convenient. It was told by someone worked at Intel in another reddit thread.
https://www.reddit.com/r/ReverseEngineering/s/hv2N41pXO5
Let's play Doom on NPU!