git/list[1] front-page[2] threads[3] people[4] search[5] about
 

Re: pack operation is thrashing my server

From
Linus Torvalds <torvalds@linux-foundation.org>
Date
Aug 14, 2008, 23:14 UTC
Message-ID
<alpine.LFD.1.10.0808141544150.3324@nehalem.linux-foundation.org>
In-Reply-To
<alpine.LFD.1.10.0808141633080.4352@xanadu.home>
On Thu, 14 Aug 2008, Nicolas Pitre wrote:
Show 6 quoted lines
> 
> Possible.  However, the fact that both the "Compressing objects" and the 
> "Writing objects" phases during a repack (without -f) together are 
> _faster_ than the "Counting objects" phase is a sign that something is 
> more significant than cache misses here, especially when tree 
> information is a small portion of the total pack data size.
Hmm. I think I may have clue.

The size of the delta cache seems to be a sensitive parameter for this thing. Not so much for the git archive, but working on the kernel tree, raising it to 1024 seems to give a 20% performance improvement. That, in turn, implies that we may be unpacking things over and over again because of bad locality wrt delta generation.

I'm not sure how easy something like that is to fix, though. We generate the object list in "recency" order for a reason, but that also happens to be the worst possible order for re-using the delta cache - by the time we get back to the next version of some tree entry, we'll have cycled through all the other trees, and blown all the caches, so we'll end up likely re-doing the whole delta chain.

So it's quite possible that what ends up happening is that some directory with a deep delta chain will basically end up unpacking the whole chain - which obviously includes inflating each delta - over and over again.

That's what the delta cache was supposed to avoid..
Looking at some call graphs, for the kernel I get:
 - process_tree() called 10 million times
 - causing parse_tree() called 479,466 times (whew, so 19 out of 20 trees 
   have already been seen and can be discarded)
 - which in turn calls read_sha1_file() (total: 588,110 times, but there's 
   a hundred thousand+ commits)
but that actually causes 
 - 588,110 cals to cache_or_unpack_entry
out of which 5,850 calls hit in the cache, and 582,260 do *not*.

IOW, the delta cache effectively never triggers because the working set is _way_ bigger than the cache, and the patterns aren't good. So since most trees are deltas, and the max delta depth is 10, the average depth is soemthing like 5, and we actually get an ugly

 - 1,637,999 calls to unpack_compressed_entry
which all results in a zlib inflate call.

So we actually have three times as many calls to inflate as we even have objects parsed, due to the delta chains on the trees (the commits almost never delta-chain at all, much less any deeper than a couple of entries).

So yeah, trees are the problem here, and yes, avoiding inflating them would help - but mainly because we do it something like four times per object on average!

Ouch. But we really can't just make the cache bigger, and the bad access patterns really are on purpose here. The delta cache was not meant for this, it was really meant for the "dig deeper into the history of a single file" kind of situation that gets very different patterns indeed.

I'll see if I can think of anything simple to avoid all this unnecessary work. But it doesn't look too good.

		Linus
Previous: Nicolas PitreNext: Björn Steinbrink
Message 52 of 80 in “pack operation is thrashing my server”
  1. Ken PrattAug 10, 2008
  2. Martin LanghoffAug 10, 2008
  3. Ken PrattAug 10, 2008
  4. Martin LanghoffAug 10, 2008
  5. Ken PrattAug 10, 2008
  6. Shawn O. PearceAug 11, 2008
  7. Ken PrattAug 11, 2008
  8. Shawn O. PearceAug 11, 2008
  9. Avery PennarunAug 11, 2008
  10. Shawn O. PearceAug 11, 2008
  11. Ken PrattAug 11, 2008
  12. Andi KleenAug 11, 2008
  13. Ken PrattAug 11, 2008
  14. Nicolas PitreAug 13, 2008
  15. Andi KleenAug 13, 2008
  16. Shawn O. PearceAug 13, 2008
  17. Shawn O. PearceAug 11, 2008
  18. Ken PrattAug 11, 2008
  19. Shawn O. PearceAug 11, 2008
  20. Andi KleenAug 11, 2008
  21. Geert BoschAug 13, 2008
  22. Shawn O. PearceAug 13, 2008
  23. Geert BoschAug 13, 2008
  24. Nicolas PitreAug 13, 2008
  25. Jakub NarebskiAug 13, 2008
  26. Shawn O. PearceAug 13, 2008
  27. David TweedAug 13, 2008
  28. Martin LanghoffAug 13, 2008
  29. David TweedAug 14, 2008
  30. Johan HerlandAug 13, 2008
  31. Ken PrattAug 13, 2008
  32. Nicolas PitreAug 13, 2008
  33. Nicolas PitreAug 13, 2008
  34. Shawn O. PearceAug 13, 2008
  35. Nicolas PitreAug 13, 2008
  36. Shawn O. PearceAug 13, 2008
  37. Nicolas PitreAug 13, 2008
  38. Shawn O. PearceAug 13, 2008
  39. Andreas EricssonAug 14, 2008
  40. Thomas RastAug 14, 2008
  41. Andreas EricssonAug 14, 2008
  42. Shawn O. PearceAug 14, 2008
  43. Nicolas PitreAug 15, 2008
  44. Nicolas PitreAug 14, 2008
  45. Linus TorvaldsAug 14, 2008
  46. Linus TorvaldsAug 14, 2008
  47. Nicolas PitreAug 14, 2008
  48. Linus TorvaldsAug 14, 2008
  49. Andi KleenAug 14, 2008
  50. Linus TorvaldsAug 15, 2008
  51. Nicolas PitreAug 14, 2008
  52. Linus TorvaldsAug 14, 2008
  53. Björn SteinbrinkAug 14, 2008
  54. Linus TorvaldsAug 15, 2008
  55. Linus TorvaldsAug 15, 2008
  56. Björn SteinbrinkAug 16, 2008
  57. Linus TorvaldsAug 16, 2008
  58. Junio C HamanoSep 7, 2008
  59. Linus TorvaldsSep 7, 2008
  60. Junio C HamanoSep 7, 2008
  61. Nicolas PitreSep 7, 2008
  62. Junio C HamanoSep 7, 2008
  63. Jon SmirlSep 7, 2008
  64. Linus TorvaldsSep 7, 2008
  65. Jon SmirlSep 7, 2008
  66. Linus TorvaldsSep 7, 2008
  67. Jon SmirlSep 7, 2008
  68. Nicolas PitreSep 7, 2008
  69. Jon SmirlSep 7, 2008
  70. Nicolas PitreSep 8, 2008
  71. Jon SmirlSep 8, 2008
  72. Jon SmirlSep 8, 2008
  73. Andreas EricssonSep 7, 2008
  74. Mike HommeySep 7, 2008
  75. Nicolas PitreAug 14, 2008
  76. Linus TorvaldsAug 14, 2008
  77. Geert BoschAug 13, 2008
  78. Dana HowAug 13, 2008
  79. Nicolas PitreAug 13, 2008
  80. Jakub NarebskiAug 13, 2008

Read the whole thread, see it on lore, or plain text.

$ cat FOOTERMessages come from the public archive at lore.kernel.org/git, fetched every hour. The front page is chosen and written each morning by an AI editor and can be wrong; the threads themselves are the record. About and API. For agents: an MCP server at https://gitlist.dev/mcp, and any thread, story or person page as Markdown by adding .md to its URL (or sending Accept: text/markdown). Details in /llms.txt.