The Exynos 9810 - Introducing Meerkat

The Exynos 9810 made a lot of noise this year as S.LSI made astounding claims of up to 2x better single-threaded performance and a 40% uplift in multi-threaded performance. We exclusively covered the first public disclosure of the microarchitecture later in January and showed that Samsung’s performance claims were not farfetched at all. Before we got back to the CPU core, let’s see what else the Exynos 9810 brings to the table.

Samsung Exynos SoCs Specifications
SoC Exynos 9810 Exynos 8895
CPU 4x Exynos M3
1c@2.7, 2c@2.3, 3-4c@1.79 GHz
4x 512KB L2
4096KB L3

4x Cortex-A55 @ 1.79 GHz
No L2
512KB L3
4x Exynos M2 @ 2.314 GHz
2048KB L2

4x Cortex-A53 @ 1.690GHz
512KB L2
GPU Mali G72MP18 Mali G71MP20
@ 546MHz
Memory
Controller
4x 16-bit CH
LPDDR4x @ 1794MHz
4x 16-bit CH
LPDDR4x @ 1794MHz

28.7GB/s B/W
Media 10bit 4K120 encode & decode
H.265/HEVC, H.264, VP9
4K120 encode & decode
H.265/HEVC, H.264, VP9
Modem Shannon Integrated LTE
(Category 18/13)

DL = 1200 Mbps
6x20MHz CA, 256-QAM

UL = 200 Mbps
2x20MHz CA, 256-QAM
Shannon 355 Integrated LTE
(Category 16/13)

DL = 1050 Mbps
5x20MHz CA, 256-QAM

UL = 150 Mbps
2x20MHz CA, 64-QAM
ISP Rear: 24MP
Front: 24MP
Dual: 16MP+16MP
Rear: 28MP
Front: 28MP
Mfc.
Process
Samsung
10nm LPP
Samsung
10nm LPE
 

At the heart of the Exynos 9810 we see four Exynos M3 CPU cores, which run at thread-count dependent maximum frequency. This ranges from up to 2.7GHz in single-threaded scenarios, to 2.3GHz in dual-core mode, and 1.79GHz in full quad-core mode.

Alongside the big performance CPUs we also see Samsung’s introduction of Cortex-A55 cores in a quad-core configuration running at up to 1.79GHz (down from the MWC units, which were running at up to 1.9GHz). The interesting thing to note is that unlike the Snapdragon 845, the A55 cores in the Exynos are in their own cluster and not shared with the M3’s.

The GPU is a new Mali G72MP18 running at up to 572MHz. The GPU configuration was a surprise here as not only did Samsung opt to go for a smaller configuration than last year’s MP20, but the clock frequency also hasn’t increased much from the Exynos 8895’s 546MHz.

On paper, the Exynos 9810 has a stronger modem than the Snapdragon 845 as it supports up to 6x carrier aggregation vs the S845’s 5xCA. The upload streams also support 256-QAM which allows for 33% higher upload speeds compared to the Exynos 8895 and Snapdragon’s modems.


Exynos 9810 Floor Plan. Image Credit TechInsights

TechInsights really delighted us this time around as they also released a die shot of the Exynos 9810 last week. The 9810 comes in at 118.94mm², which is 14% bigger than the 8895’s 103.64mm². Qualcomm seems to have an edge in total die size and we don’t have to look very closely to notice why.

At 20.23mm² the Exynos M3 complex is absolutely massive compared to other mobile SoC CPUs. At 3.46mm² for the core and accompanying L2 the Meerkat core is over twice as big as the 1.57mm² of the A75+L2 in the Snapdragon 845, granted that the latter has half the L2 cache. Meerkat indeed almost matches Apple’s Monsoon cores in the A11 which come in at 2.68mm² - but only if one takes into account the L2 cache of the M3 for which I roughly estimate 0.88mm². Apple also has a slight density advantage due to TSMC’s 10FF manufacturing node.

Nevertheless, at a total of 22.1mm² for both clusters Samsung has thrown down a lot of silicon for the CPU complexes, far more than Apple’s 14.48mm² and Qualcomm’s 11.39mm².

An interesting aspect we can see in the die shot is the way that Samsung arranges the L3. Indeed we reached out again to ARM for clarification on if the DSU allows third-party cores to be used, and contrary to we had been told last year, ARM doesn’t enable third-party cores to be connected. This means that the L3 we see here on the M3’s are of Samsung’s own design. At 4MB, the cache is quite big, but as mentioned, the layout is unlike anything we’ve seen before as it looks like not only does Samsung distribute the L3 SRAM banks in a row/column, but the L3 arbitration logic and L3 tags are also distributed among the banks alongside each M3 core.

Image Credit TechInsights

Looking closer at the core we see two L2 banks (512KB total) along with their tag buffers on the left side. At the top middle we see what is likely the 64KB L1D cache with the load/store engine. On the right side, likely on the bottom, we see the L1I cache as well as other front-end related memories. Unfortunately Samsung’s physical implementation here is a sea of gates and it’s hard to make out the individual CPU engines.

The Cortex-A55 cluster looks quite similar to the Exynos 8895’s A53 cluster – this is due to the lack of per-core L2’s and only a shared 512KB L3 that essentially acts as a shared L2. The performance degradation of the A55s not having L2’s is offset somewhat by the fact that the L3 is run at the same frequencies as the cores – eliminating the need for asynchronous bridges between the cores and the cache and thus reducing the L3 cache latency compared to a normal DSU configuration.

Finally the last interesting take-away from the die shot is the GPU. The Mali G72MP18 comes in at a total of 24.53mm² which is smaller than last year’s >~32mm² behemoth based on the Mali G71MP20. Here it’s clear just much of an advantage Qualcomm has as the Adreno 630 with its 10.69mm² is outright tiny compared to the Mali and even has a significant lead even over Apple’s A11 GPU which comes at 15.28mm².

2.9GHz.. 2.7GHz .. 2.3GHz ... 1.79GHz?? Which is it?

One of the bigger discussion points about the Exynos 9810 was its clock frequency. Samsung had initially announced a peak clock frequency of 2.9GHz but immediately that seemed unlikely given S.LSI’s history of backing down on their initial frequency claims.

Looking at the voltage tables of the Exynos 9810 points out quite a wide range of voltages for the M3 cores, but it’s the far end that seems quite problematic. To actually reach the initially advertised state of 2.9GHz (2860MHz), it takes an extremely high voltage of 1213mV. Backing down to 2704MHz with which the Galaxy S9 is released ends up with an already drastic decrease of 100mV. Historically Samsung has always had quite high voltages at the far end of the frequency tables as it seems they optimise the physical implementation for leakage and power, which in turn requires higher voltages to reach high frequencies.

When looking at the power curves correlated with our traditional integer power virus we see that there’s an immense increase in power consumption at the higher frequencies. Indeed going from 2.3GHz to 2.9GHz would have doubled power usage, and even 2.7GHz comes at a steep power price. Given that power usage scales roughly along the lines of voltage cubed, the SoC's efficiency suffers with the increased frequency. The good news here is that Samsung’s efficiency curve is quite steep and linear, that means backing down on frequency should see significant efficiency gains.

Samsung’s decision to limit 2+ core frequencies makes sense in the context of thermal constraints. Even if a CPU core is very efficient at its peak performance, it’s just physically not possible to run multiple cores at peak performance as the SoC just lacks the required thermal dissipation. It’s also important to emphasise this difference between power usage and efficiency: This is no Snapdragon 810 situation where we have high power but with lacking performance. Total energy usage of the M3 should thus be equal to a lower performance core which uses less power.

The only comment I’d like to add here is that I think Samsung would have done a lot better if the M3 cores had been split into a 2+2 configuration with separate frequency and voltage planes. This is something we'll get back to in the battery life section of the review.

I’ve had a look through Samsung’s scheduler and DVFS mechanisms which controls the switching between the 1/2/3/4 core modes and generally I’ve been unimpressed by the implementation. Samsung had made use of hot-plugging to force thread migrations between the cores which is an inefficient way of implementing the required mechanism. The scheduler is also tuned extremely conservatively when it comes to scaling up performance, also something we’ll see the effects of in the system performance benchmarks.

Lastly, I noticed that the commercial unit we acquired had quite different DVFS settings that I was unable to reproduce the excellent memory latency scores I had measured on the launch event devices at MWC. This means that the memory performance is going to be less than I had anticipated, and metrics such as full random access latency actually saw a degradation compared to the Exynos 8895.

The Snapdragon 845 - A Quick Recap CPU Battle - SPEC Performance & Efficiency
Comments Locked

190 Comments

View All Comments

  • jospoortvliet - Monday, March 26, 2018 - link

    Thou did an amazing job explaining/analysing why the Exynos performs so bad in practice for a SOC with such powerful cores. THANK YOU for that as it has been missing from every other test, leaving readers speculating. I hope SAMSUNG can address these problems and shame on them for releasing a devise with such a bad tuning! Makes them look bad and incompetent, honestly, their software/kernel team needs a real kick in the but for making the CPU design team look so bad! Perhaps indeed the M3 hasn't caught up to the efficiency of Qualcomm but the (relatively) horrible performance in tests is certainly not due to bad SOC design...
  • tvdang7 - Monday, March 26, 2018 - link

    LTE battery test?
  • SirCanealot - Tuesday, March 27, 2018 - link

    No idea if this is current information, but Andrei has started in the past that he has really spotty reception where he is based; so I believe it is difficult for him to test battery life for cellular connections. Please correct me if I am wrong :)
  • Andrei Frumusanu - Wednesday, March 28, 2018 - link

    We've learned a lot about cellular testing and the sheer amount of variables that aren't controllable are crazy. For one I found out the carrier that I was on didn't have CDRX enabled on its base stations which worsened power usage by up to 30% on the handset. Since then I have a different carrier and better signal, but there's still questions about how representative the results are.

    For this review I didn't have time to test all devices since I have to do it sequentially while for Wi-Fi testing I can have all devices running at the same time.
  • robertkoa - Monday, July 9, 2018 - link

    Yes. On my recently acquired [ 2 weeks ] I get very long Screen On Times [Qualcomm USAversion] in large part I assume to :
    ALWAYS have 2 to 5 bars of 4GLTE Signal from Metro PCS/ TMobile Towers- with 4 and 5 Bars inside large Malls (!). [ Phone NEVER changes Bands on Display ].
    ■Also - indoors I like low screen brightness. 30%
    ■I have very few APPS that are allowed to send Notifications.
    ■No SITES can send Notifications.
    ■All/ Most Options are OFF or asleep till I need them - Location/Sync/ Wifi/ etc etc etc
    ■ I use the Battery Maintenance "OPTIMIZE" a few times a day - there is no downside - Apps it puts to sleep come right back.
    ■ Apps don't choose to turn on things or use Battery as much as I can stop them ...lol.

    First Nonremovable Battery Device....so a little overvigilant right now.
    INCLUDING: not sure IF I will dowload HUGE 750 meg Update File of 'security updates' - I feel it might include constant background running processes and a larger Operating System and cost me battery life.
    I can browse on 4G LTE 8 to 10 hours with 8 to 10 hours SOT now.
    With an hour of calls and 45 minutes Youtube included ...7 to 8 hours SOT.
  • pjcamp - Monday, March 26, 2018 - link

    Does it have a UI? Because I see no evidence of that in the index to this review. I haven't owned a Samsung since the S3 -- in Touchwiz still a shi77y thing?
  • babadivad - Friday, April 13, 2018 - link

    "Does it have a UI??"

    You could't interact with it at all if it didn't have a UI.
  • tipoo - Monday, March 26, 2018 - link

    A little interesting that the four 'big' cores are smaller than the four 'little' cores, such that they would have used less die space with 8 of the gold. But I assume that's because more transistors were spent making X performance level more efficient.
  • tipoo - Monday, March 26, 2018 - link

    Also interesting how much we're still learning anew about the iPhone X display and how much Apple thought of that we still hadn't puzzled out, that the processing is very different than Samsungs despite being made by them.
  • id4andrei - Monday, March 26, 2018 - link

    Samsung "cobbled" color profiles to Android in an effort to provide some color management. Ios is still the only mobile os with a proper color management system. Until Google does the same, ios will be better, even if the iphone has lcd.

Log in

Don't have an account? Sign up now