Log in

View Full Version : Evaluation of HEVC decoders (SW, Hybrid and HW)


Pages : 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 [40] 41 42 43 44 45 46 47 48 49 50 51 52

peca89
25th July 2017, 12:17
Hello,

I'm following this forum for a month now and I didn't find definitive answer if a new GT 1030 would be adequate upgrade for my years old HTPC. It will be connected directly to 4K HDR TV using HDMI.


Can it upscale 1080p60 and output 2160p60 using EVR? Can it do it using madVR?
Can it decode AVC/HEVC/VP9 2160p60 SDR and output 2160p60 SDR? EVR or madVR?
Can it decode AVC/HEVC/VP9 2160p60 HDR and just pass it through to a HDR compatible TV, therefore skipping resource-hungry 4K scaling and HDR<->SDR conversion?

NikosD
25th July 2017, 12:22
All these are good questions, but I'm afraid GT 1030 is not so popular so the one and only owner appeared here has declared that he doesn't own a 4K display.

We must find a combination of 1030 using a 4K display, which could be rare considering low GPU power of the card and only 2GB VRAM.

jmonier
25th July 2017, 13:58
Based on my limited experience with a GTX1060 3gb, I would say there's no chance. In all these cases, memory usage was well above 2 gb and seemed to be pushing the limits for both memory and performance. I'm going to a GTX1070 as a result.

v0lt
25th July 2017, 19:48
@NikosD
2 GB is enough to play 4k video. Of course, you can specially tweak the settings, but this is not an indication that the video card is bad.

NikosD
25th July 2017, 20:03
The problem could be the output of 4K not the decoding and then downscaling to 1080p.

I don't have a 4K display, but seems strange that Nvidia has set the requirement to >3GB for 4K Netflix playback, if 2GB could be enough.

huhn
25th July 2017, 20:04
while it is possible to playback UHD videos at UHD with 2GB Vram you are really pushing it.

i was using a 960 with 2Gb for sometime doing this just opening a web browser was pushing it over the limited for me.

cutting corners by using 8 bit buffers like mpc-BE may help in term of Vram usages but that doesn't mean everyone is unable to see the banding...

NikosD
25th July 2017, 23:02
Good news!
It seems that latest Chrome Canary v62 (x64) works like a charm using VP9 hybrid acceleration of a Polaris card and Creators Update!

Firstly, like plain Chrome v59 (x64) it recognizes out of the box VP9 Youtube content but it doesn't enable by default the hybrid VP9 acceleration.

BUT if you force enable VP9 acceleration using the --enable-accelerated-vpx-decode=2 switch, you see no stuttering in the video decoding and it has a minimal frame dropping of just 35 frames during the whole duration of the very difficult 4K60 fps Youtube clip posted above.

The GPU load works exactly like Chrome, but without stuttering and you see the GPU and memory clocks going to the highest levels.
The CPU usage is around ~50% on average.

So, it seems that AMD's VP9 hybrid acceleration works, at least on one Chromium based browser and force enabled.

Of course Canary has some strange behavior as an experimenting browser, so we definitely need a more stable overall browser like Chrome to implement correctly the VP9 hybrid acceleration.

wanezhiling
26th July 2017, 01:40
https://techreport.com/news/31627/amd-publishes-patches-for-vega-support-on-linux
The UVD 7.0 decoder and Video Coding Engine 4.0 encoder are reported to be included in the upcoming Vega based GPUs.
Hope AMD could catch GM206 now..

v0lt
26th July 2017, 04:28
@NikosD
Netflix wrote a complete nonsense. "GTX 1050 (not Ti) with 3 GB of video memory".

NikosD
26th July 2017, 05:23
No. It was not Netflix.

I told you what Nvidia wrote in their requirements.

>3GB VRAM

el Filou
26th July 2017, 20:50
All these are good questions, but I'm afraid GT 1030 is not so popular so the one and only owner appeared here has declared that he doesn't own a 4K display.

We must find a combination of 1030 using a 4K display, which could be rare considering low GPU power of the card and only 2GB VRAM.
Spec-wise, a 1030 is pretty much half a 1050 Ti in every way: half the quantity of VRAM, half the number of transistors and half the number of compute cores, and nearly half (43%) the memory bandwidth.
So in theory you just need someone with a 1050 Ti and a 4K TV (sorry, 1080p for me), and ask them to watch with GPU-Z if GPU usage, VRAM usage, and memory controller usage never go above 50% (or let's say 40-45% to keep a safe margin) and if it doesn't then it should be fine for a 1030, right?

huhn
27th July 2017, 00:34
you don't need a 4k display to test 4k "output" at least not with nvidia.

the problem is the Vram and nothing else. if you do a proper render chain with proper buffer you will run out of Vram or close to it.

peca89
28th July 2017, 08:06
Thanks for your replies. I must admit I do not really understand why do you think (or why it is) 4k so demanding. If you accept the logic that performance and ability to play video in realtime can be linearly scaled like some people here mentioned, can you explain why my ages old Radeon 5670 with only 512MB of VRAM and only 400 MHz of clock during playback is perfectly capable of chewing 1080p at 60 Hz. It stays below 250 MB of memory usage and below 50% of GPU usage using Win 10, MPC-HC, EVR and LAV video with DXVA native. 4k is "only" 4 times more pixels than 1080p and 1030 has 4x more VRAM, 3x GPU clock and much faster shaders.

Main question: Why does 4k need >2GB of VRAM when 1080p needs 10 times less?

el Filou
28th July 2017, 15:01
Thanks for your replies. I must admit I do not really understand why do you think (or why it is) 4k so demanding. If you accept the logic that performance and ability to play video in realtime can be linearly scaled like some people here mentioned, can you explain why my ages old Radeon 5670 with only 512MB of VRAM and only 400 MHz of clock during playback is perfectly capable of chewing 1080p at 60 Hz. It stays below 250 MB of memory usage and below 50% of GPU usage using Win 10, MPC-HC, EVR and LAV video with DXVA native. 4k is "only" 4 times more pixels than 1080p and 1030 has 4x more VRAM, 3x GPU clock and much faster shaders.

Main question: Why does 4k need >2GB of VRAM when 1080p needs 10 times less?
You are right, and 2 GB is well enough for "basic" 4K video playback (Edit: except for Netflix, if you believe NVIDIA, but they've never explained the exact reason), it's just that quality is a highly subjective question.
A lot of people on this forum are videophiles, and if you want to take full advantage of the processing power of a PC GPU to enhance the rendering quality and smoothness of your video playback, then you'll have to use things like a long buffer queue, internal processing using high precision floating point formats etc. All of this combined can increase VRAM usage quite a lot.

Here are some numbers on my 1050 Ti, playing HEVC Main10 2160p60 140 Mbps (i.e. Blu-ray UHD level) with DXVA Native, and outputting to 1080p:

MPC-HC EVR: 705 MB
MediaPortal EVR-CP: 1110 MB
MPC-HC madVR: 1483 MB
MPC-HC EVR-CP 8 buffers: 1765 MB (I can't understand why it uses more than madVR)
MediaPortal madVR: 1802 MB

Using DXVA CopyBack instead of Native adds 370 MB to those numbers.
Outputting to 2160p 10 bit instead of 1080p would add a few 100 MBs again.

v0lt
28th July 2017, 16:03
MPC-HC EVR-CP 8 buffers: 1765 MB (I can't understand why it uses more than madVR)
Why 8 buffers? What is the format of these buffers?

el Filou
28th July 2017, 17:56
Why 8 buffers? What is the format of these buffers?
The renderer setting is just called "EVR Buffers", the default is 5 but I found that increasing that value a bit prevents skipped frames in rare conditions.
I guess it's a setting equivalent to madVR's "GPU queue size".

huhn
28th July 2017, 19:46
@el Filou

you are aware that at UHD output the render queue and the output frame are at least 4X as big which is more than 100 mb...

like i said before you can test it using nvidia DSR. and again UHD output with 2 GB is possible but you are really pushing it.

el Filou
28th July 2017, 22:23
@el Filou you are aware that at UHD output the render queue and the output frame are at least 4X as bigYes I am.

which is more than 100 mb...I know, that's why I wrote "a few 100 MBs", but I should have written "a few hundreds".

like i said before you can test it using nvidia DSR. and again UHD output with 2 GB is possible but you are really pushing it.Thanks for the clarification, I didn't understand how it could be done when you said we could test 4K output.
I've tried DSR and here is the VRAM usage I find in 2160p, with DXVA Native:

EVR: 1105 MB
EVR-CP, render queue 5: 1775 MB
madVR (my custom settings): 2480 MB :o
madVR shortest render queue: 1988 MB.

So EVR is fine, MPC-HC's EVR-CP is still alright but approaching the limit.

huhn
28th July 2017, 23:34
windows high DPI settings can take up some Vram just opening a web browser can make it harder too.

adding needed 10 bit processing and output is going to make it even harder.
subtitles and other things too.

i made madVR work with 2 GB too but i will never ever recommend someone to do the same so do yourself a favor and get 4 Gb Vram or more.

v0lt
29th July 2017, 05:25
Never use 32-bit Floating Point textures (in MPC-HC, this is called Full Floating Point Processing). This leads to a meaningless consumption of video memory.

If you really need a floating point for intermediate results, then use 16-bit Floating Point textures (in MPC-HC, this is called Half Floating Point Processing). This is quite enough.

I'll advise you to see the video settings and playback statistics in MPC-BE. There it is more clear.

edwdevel
29th July 2017, 17:42
Hello,

I'm following this forum for a month now and I didn't find definitive answer if a new GT 1030 would be adequate upgrade for my years old HTPC. It will be connected directly to 4K HDR TV using HDMI.


Can it upscale 1080p60 and output 2160p60 using EVR? Can it do it using madVR?
Can it decode AVC/HEVC/VP9 2160p60 SDR and output 2160p60 SDR? EVR or madVR?
Can it decode AVC/HEVC/VP9 2160p60 HDR and just pass it through to a HDR compatible TV, therefore skipping resource-hungry 4K scaling and HDR<->SDR conversion?


The short answer to your questions is: if you want to use madvr with a 4k TV,
you will probably not be happy with a GT 1030. :)

The long answers to your questions are:

First, I am the poster here who benchmarked several clips using a GT 1030 and a
wolfdale 2core 3Ghz and a 1600x900 monitor. Since I posted the benchmarks,
my old FHD LCD TV died, and has since been replaced with a brand new Sony XBR900E UHD TV.
Unfortunately, I don't have a HTPC (nor will I), so I still won't be able to
definitively tell you that a 1030 can drive a 4k TV. However, I have been doing
some research with madvr vs. EVR using the 1030, and I believe that some reasonably
confident conclusions can be made about the performance of the GPU with 4k output.

Several posters here have shown that the 1030 with its' 2GB memory *should* be able
to drive a 4k TV. Also, is it obvious that there is nothing special about the 1030 in
Nvidia 10 series line of 1050, 1060, etc. when it comes to 4k output? Nvidia has not
come out with a bulletin stating that, sorry, our 1030 just cannot drive 4k output :).
I have done playback benchmarks using DXVA Checker with native 4k resolution and EVR
with various 4k clips. In "Colors of Journey" I got over 100fps, with low cpu and 100%
GPU (this may be normal for playback benchmarks). It is illogical to think that
the GPU scanning electronics cannot drive an HDMI output at 4K from a video memory
resident frame. Thus, I think we can confidently state that the 1030 can display 4k60fps
clips using HEVC and VP9 and EVR.

Madvr, though, is a completly different story. I benchmarked a 4k24fps clip "Exodus"
downsizing to monitor size (about 1080p) with MPC-BE and madvr, and basically
got perfect beautiful HDR output, with the cpu at 10% and the gpu at 90%- without DXVA
downsizing. As soon as DXVA downsizing is set on, the GPU goes to 20% (though the
colors are all wrong- I reported this as a bug on the madvr forum here). This
demonstrates the real issue with madvr: it basically is an image processor using D3D
GPU function for even simple downsizing and is completely dependant on the performance
of the GPU board. My 1030 cannot play 4k60fps clips with madvr bilinear downsizing
to 1080p- not because of the 2G memory, but simply because the 1030, which is the
weakest GPU in the 10 series, cannot keep up with madvr's demands. I can say, though,
both "Exodus" and "Life of Pi" 4k24fps can play perfectly- downsized to 1080p.

So, if you want to use madvr to do image processing such as downscaling, upscaling,
sharping or basically anything, you can forget about using the 1030. You should hang
out in the madvr forum here, where people complain about $600 GTX 1080 boards having
performance problems. :)

But why do you even need an HTPC if you are going to a modern 4k TV? I've been enjoying
my new android wi-fi state of the art full array dimming UHD TV for the past few days,
which can play all of my 4k clips using wi-fi DLNA perfectly, not to mention all 4k
youtube clips and, yes, Netflix UHD, directly over wi-fi connected to the router (though
I did have problems with the old wi-fi speed, and have since upgraded both my router and
internet connection- but that is another story)

Regards,
edwdevel

NikosD
29th July 2017, 19:35
Major performance drop using latest 17.7.2 drivers.

The GPU clocks have been stabilized and using pure decoding mode, the memory clock is always at min state (300 MHz) and the GPU clock is at min state (300 MHz) around 90% of the whole decoding process time, going up to 466 MHz occasionally.

Of course pure video decoding mode has nothing to do with GPU clocks, but memory clock always at min state (300 MHz) could be a problem.

AMD has probably put the internal video decoder clocks lower (they are invisible to the user)

The result is a pure decoding speed less than the <17.4 drivers, a lot less to be more exact. Huge drop in performance.

Now, during playback benchmark mode we are on the opposite situation, regarding GPU/ memory clocks.

GPU clock and Memory clock are at max speed, but still the decoding speed is less than previous drivers, but better than <17.4 drivers.

I also got a micro-stuttering issue while moving cursor on the desktop, it seems that the cursor pauses just for a sec with no reason, like some older situations when incompatible drivers have been installed on the system or some other kind of interrupt caused that little pause.

17.7.2 is a major update, but a probably a little unpolished.

I don't know if the decision of slowing down internal video clocks of the video decoder (invisible to the user) in pure decode mode is permanent for AMD.

el Filou
30th July 2017, 10:56
Major performance drop using latest 17.7.2 drivers.

The GPU clocks have been stabilized and using pure decoding mode, the memory clock is always at min state (300 MHz) and the GPU clock is at min state (300 MHz) around 90% of the whole decoding process time, going up to 466 MHz occasionally.

AMD has probably put the internal video decoder clocks lower (they are invisible to the user)Did they slow it down so much that you can't play 2160p60 smoothly?

NikosD
30th July 2017, 11:01
In playback mode the decoding speed is:

Drivers earlier than 17.4 < 17.7.2 < Drivers after 17.4.x up to 17.7.1

So, you can play 2160p60 smoothly, but previous drivers from 17.4.x up to 17.7.1, were definitely faster.

el Filou
30th July 2017, 11:19
Maybe they introduced a bug while trying to optimise power profile?
I understand that in this new computing landscape where laptops and other mobile devices represent the mass market, you'd want the lowest possible power consumption while still providing real-time playback, but they should provide an advanced option to put the video decoder to highest performance state for desktop parts.

aufkrawall
30th July 2017, 15:22
I've followed Phoronix news quite thoroughly during the last months, and I think there's still not a single sign of AMD supporting VP9 with Vega on Linux. :(
Nvidia's cuda decoding isn't great either since it has more often problems with certain formats (e.g. HEVC 10 bit HDR), but at least it works well for VP9 YT videos (and without quality degradation, unlike shit APIs by Microsoft).

nevcairiel
30th July 2017, 17:33
it works well for VP9 YT videos (and without quality degradation, unlike shit APIs by Microsoft).

Microsofts decoding APIs are perfectly bit-exact decoding. Try not to spread false rumors.

NikosD
30th July 2017, 19:35
Least expensive RX VEGA 56 will cost around 400$, so I don't think a lot of people would buy it for HTPC or watching YouTube videos.

Only video quality freaks using madVR could find that card interesting for video playback, along with 3D gaming of course.

But I think the plans of AMD involve a desktop APU release on early 2018, along with some medium level cards based on VEGA replacing Polaris cards, so yes, in a longer term it's important for VEGA core to support VP9 hardware decoding.

aufkrawall
30th July 2017, 21:53
Microsofts decoding APIs are perfectly bit-exact decoding.

You likely know the issues very well, no need to be extra-smart.

nevcairiel
30th July 2017, 22:39
You likely know the issues very well, no need to be extra-smart.

In which case you likely do not, otherwise you would know better.

There are absolutely no issues to get a perfect bit-exact image out of the hardware decoder using DXVA. FFmpeg has regression tests that test the bit-exactness of decoding, you are free to run those with DXVA decoding to confirm this fact.

I can only guess you are referring to the DXVA-native "issue" that madVR has (and some other players following a similar implementation).
This is not related to decoding, the decoder gives you a perfect bit-exact image back, it only gets degraded through the process of how these renderers convert it into a texture - which is a D3D9 limitation (lack of 4:2:0 textures). You could use DXVA with D3D11 to avoid this (ffmpeg calls this "d3d11va"), or even map the decoded surface to OpenGL - or the easy alternative, albeit with a performance hit (although in my experience on NVIDIA the impact is barely measurable), use Copy-Back.

If you build against Direct3D, using D3D11 is clearly the best option, because you should be using that anyway to get NV12/4:2:0 support for your textures, so you don't have to cheat the chroma upsampling shaders.

NVDEC decoding (formerly CUVID) isn't any different. Either you use copy-back or you map the CUDA memory as a OpenGL surface, or even a D3D surface - although if you use D3D9, the same problems would manifest themself.

With any decoding API one of the main problems is how to get the image from the decoder format (typically a simple GPU memory buffer) into a format for display (typically a texture), but this is generally more limited by the display API you are using, not the decoder. In this case above, D3D9 (not DXVA) is just limited, but we have a modern replacement already.

aufkrawall
30th July 2017, 23:06
Which renderer does not show quality issues in shape of chroma blur and additional banding on Nvidia DXVA2/d3d11va without copyback?
madVR, MPDN and mpv do. I also gave Microsoft's Windows 10 video app a try some time ago, looked bad too (might have to recheck this, but not expecting anything).
That there is no quality degradation by just outputting the frame to memory is a totally different topic which I wasn't talking about. Read more carefully perhaps?
And mpv also has the issue with d3d11va and DX11 ANGLE renderer.

The issues do not occur with native cuda (no copyback) on both Windows and Linux and neither with VDPAU (no copyback) on Linux.
It's only bad with Microsoft APIs involved, such coincidence.
(The sames goes btw. for GPU compositing in browsers.)

nevcairiel
30th July 2017, 23:15
Its perfectly possible to do this with D3D11 (which madVR or MPDN cannot do right now, since they don't have a decoder yet that outputs D3D11 surfaces - in fact does MPDN support DXVA-native at all even?), or even OpenGL (in theory, I haven't personally tested the DXVA -> OpenGL interop)
I don't know what mpv does in d3d11va zero-copy mode, but considering its primarily linux-focused, I wouldn't expect this to be a fully optimized path.

I wouldn't expect the Microsoft Video App to behave high-quality under any definition, it just tasks the GPU to convert the video to RGB, it has no interest in getting a bit-exact copy of the decoded video. The same thing browsers do.


That there is no quality degradation by just outputting the frame to memory is a totally different topic which I wasn't talking about. Read more carefully perhaps?


You didn't actually specify any topic you were talking about, just generically referenced "decoding".

Since you seem to have tested that quite extensively, I would be interested in a sample that shows these differences quite obviously for future reference.

NikosD
31st July 2017, 06:56
VEGA supports "Trusted memory zone" for Ultra HD Blu-ray and Netflix 4K:

http://www.tomshardware.com/news/amd-radeon-rx-vega-64-specs-availability,35112.html

aufkrawall
31st July 2017, 20:02
Since you seem to have tested that quite extensively, I would be interested in a sample that shows these differences quite obviously for future reference.
It's true that, unlike direct DXVA2, direct d3d11va doesn't show quality degradation with 8 bit input and the Windows 10 video app can also play HEVC 10 bit video without banding (just rechecked it).
However, mpv is limited to rounding down 10 to 8 bit with d3d11va, so probably the decoder can't be told to dither down to 8 bit and angle can't deal with 10 bit input.
Quite stupid limitation, not a problem with cuda or vdpau (and I guess neither vaapi).

wanezhiling
1st August 2017, 23:38
Nothing special, just add what Apollo Lake lack

NikosD
4th August 2017, 09:06
Another CPU launch, another fraud from Intel.

Coffee Lake is incompatible with Kabylake's chipset, but Intel had said that it will be compatible.

All the happy owners of 100 and 200 series are moving to Ryzen, the best overall mainstream processor ever.

Nobody with a 6C/12T CPU is interested in iGPU of Coffee.

Not even the biggest hard core Intel fan boys like you are going to but that useless and expensive CPU when Ryzen 7 has 8C/16T and is cheaper.

huhn
4th August 2017, 09:33
to bad that most people that doesn't really need a PC right now are avoiding to buy one.

nevcairiel
4th August 2017, 09:50
Another CPU launch, another fraud from Intel.

Coffee Lake is incompatible with Kabylake's chipset, but Intel had said that it will be compatible.

I don't remember them saying that anywhere officially, and none of the articles talking about the Coffee Lake and 200-series incompatibility have mentioned anything of that sort, got any source for that?

Not really interested in buying a Coffee Lake either way, though, but rumors have a 4c/8t i3 showing up, which might be good value (first quad-core i3, if its true)

NikosD
4th August 2017, 11:33
LGA 1151 is LGA 1151 whether it's Skylake, Kabylake or Coffee Lake.

The entire planet thought that Intel will eventually change its notorious mind, but still they keep going on the same way.

6C/12T for mainstream desktop and i3 with 4C/8T (if true) are the consequences of Ryzen competition.

huhn
4th August 2017, 11:47
that's not how a chipset works...

peca89
9th August 2017, 09:11
I've given up waiting for relevant information to show up on somewhere on the internet so I decided to buy the GT 1030 and test it myself. Let me first answer my own questions:


Can it upscale 1080p60 and output 2160p60 using EVR? Yes, absolutely!
Can it do it using madVR? Did not test properly yet.
Can it decode AVC/HEVC/VP9 2160p60 SDR and output 2160p60 SDR? Yes, absolutely! No problem at all.
EVR or madVR? Tried only EVR so far
Can it decode AVC/HEVC/VP9 2160p60 HDR and just pass it through to a HDR compatible TV, therefore skipping resource-hungry 4K scaling and HDR<->SDR conversion? I managed to set up madVR to trigger HDR mode on TV. However, colors are completely wrong in fullscreen. I need more time figuring out how madVR works


Everything I threw to the card (up to 4K res), it managed to decode and render properly. I've tried most of the HEVC test clips mentioned in this topic and everything works fine except higher than 4K resolutions. VRAM usage at 4K 60fps 10bit HEVC output at 4K@60Hz is up to 1500MB when using EVR.

If anyone has an idea how to set up madVR to do basically nothing and just pass HDR through, please write here. I know that's not what madVR was made for, but it is probably only thing possible using low end video card. Also, I can do some more tests if someone is interested.

huhn
9th August 2017, 09:21
don't use DXVA native with madVR.
set the GPU queue to 6 and pray.
don't use FSE with HDR for know it's kind of buggy in the nvidia GPU driver.

give the windows 10 HDR API a try with 10 bit windowed.

it is not possible to do basically nothing to an decoded image using madVR or EVR.

try either DXVA chroma scaling or bilinear both are absolutely terrible...

and make sure no other program like a web browser is running in the background.

jmonier
9th August 2017, 13:09
For me, the few HDR samples that are available (Life of Pi, Exodus, etc.) work perfectly with FSE under Win 8.1 and the latest Nvidia driver.

nsnhd
9th August 2017, 13:13
Can it decode AVC/HEVC/VP9 2160p60 HDR and just pass it through to a HDR compatible TV, therefore skipping resource-hungry 4K scaling and HDR<->SDR conversion?

Also, I can do some more tests if someone is interested.

For those who don't own HDR displays, HDR ->SDR conversion is crucial. Could you pls test that to confirm the GT 1030 capability ? Thanks

peca89
9th August 2017, 20:43
Thanks for the tips. I need to reinstall Windows to get Creators Update to get Windows HDR API because I'm currently on LTSB 2016...

HDR->SDR conversion is mixed success. The best I could get is successful 4K conversion at 24fps, but only if I set "preserve hue in low quality" in madVR settings. Life of Pi sample plays fine. All of 60fps 4K HDR clips I have (e.g. Sony camp) play as a slideshow. 1080p HDR samples at 60fps in VP9.2 (downloaded from Youtube) play fine.

Since madVR requires copy-back (thanks huhn) it's no surprise that 4K 60fps does not work. Even SDR 4K 60fps using EVR does not work if copy-back is used (memory gets maxed out).

But hey, 24fps is fine for movies! :)

huhn
10th August 2017, 05:02
madVR works with DXVA native but with native it is using more Vram. i guess the decode queue is in the Vram too now.

NikosD
12th August 2017, 08:05
Intel hasn't unveiled a new architecture since August 2015, which was Skylake.
It's been already two years.

The "new" Coffee Lake CPUs are actually clones of AMD's concept of multi-core.

So, probably yes, Core i3 8300 could be a 4C/8T chip, matching all previous generations Core i7s from 2011 Sandybridge at least, regarding number of cores/threads.

Digging in Intel's internals, it seems they face a serious problem of not being able to produce desktop powered/ performance chips using 10nm process - only low-powered.

Intel has a broken 10nm process right now.

With Ryzen/Threadripper/EPYC chips alive and kicking, things couldn't be worse for Intel this period.

Let's see their funny reactions, like Coffee Lake for example.

el Filou
12th August 2017, 15:47
Digging in Intel's internals, it seems they face a serious problem of not being able to produce desktop powered/ performance chips using 10nm process - only low-powered.

Intel has a broken 10nm process right now.
Who else has tried to produce 4 GHz+ high power chips in 10 nm? Samsung and TSMC have only produced mobile chips in 10 nm until now, and Samsung '10 nm' is barely finer than Intel's 14 nm, so Intel has quite a margin of development IMO.
Let's wait and see how (AMD) - sorry, GloFo - manages 10 nm too.

NikosD
12th August 2017, 15:52
GloFo is moving directly to 7nm for Zen 2 and Navi GPU.

nevcairiel
12th August 2017, 15:55
Comparing the marketing process node names is impossible at the current time. Companies have been using different references to measure them, which leads to some smaller numbers that don't really indicate an advantage. Intel may be playing catchup on their lineup, but their processes are still quite stable.
As el Filou already said, Samsung 10nm isn't really smaller then Intel 14nm, its just a different reference of measurement. The same can be applied to TSMC as well. Don't know the details about GloFo off-hand.