server-skynet-source-3rd-jemalloc

project-base/server-skynet-source-3rd-jemalloc

Author	SHA1	Message	Date
David Goldblatt	aea91b8c33	Clean up some minor data structure inconsistencies Namely, unify the include guard styling with the majority of the project, and do flat_bitmap -> fb, to match its naming convention.	2021-05-12 11:14:23 -07:00
David Goldblatt	1f688490e1	Stats: Fix a printing bug when hpa_dirty_mult = -1 Missed a layer of indirection.	2021-05-05 19:45:25 -07:00
David Goldblatt	4f7cb3a413	Sized deallocation: fix a typo. dealloction -> deallocation.	2021-05-04 16:46:15 -07:00
Qi Wang	9b523c6c15	Refactor the locking in extent_recycle(). Hold the ecache lock across extent_recycle_extract() and extent_recycle_split(), so that the extent_deactivate after split can avoid re-take the ecache mutex.	2021-03-31 14:42:33 -07:00
Qi Wang	ce68f326b0	Avoid the release & re-acquire of the ecache locks around the merge hook.	2021-03-31 14:42:33 -07:00
Qi Wang	7dc77527ba	Delete the mutex_pool module.	2021-03-29 17:19:53 -07:00
Qi Wang	03d95cba88	Remove the unnecessary arena_ind_set in base_alloc_edata(). All edata alloc sites are already followed with proper edata_init().	2021-03-29 17:19:53 -07:00
Qi Wang	3093d9455e	Move the edata mergeability related functions to extent.h.	2021-03-29 17:19:53 -07:00
Qi Wang	7c964b0352	Add rtree_write_range(): writing the same content to multiple leaf elements. Apply to emap_(de)register_interior which became noticeable in perf profiles.	2021-03-29 17:19:53 -07:00
Qi Wang	add636596a	Stop checking head state in the merge hook. Now that all merging go through try_acquire_edata_neighbor, the mergeablility checks (including head state checking) are done before reaching the merge hook. In other words, merge hook will never be called if the head state doesn't agree.	2021-03-29 17:19:53 -07:00
Qi Wang	49b7d7f0a4	Passing down the original edata on the expand path. Instead of passing down the new_addr, pass down the active edata which allows us to always use a neighbor-acquiring semantic. In other words, this tells us both the original edata and neighbor address. With this change, only neighbors of a "known" edata can be acquired, i.e. acquiring an edata based on an arbitrary address isn't possible anymore.	2021-03-29 17:19:53 -07:00
Qi Wang	1784939688	Use rtree tracked states to protect edata outside of ecache locks. This avoids the addr-based mutexes (i.e. the mutex_pool), and instead relies on the metadata tracked in rtree leaf: the head state and extent_state. Before trying to access the neighbor edata (e.g. for coalescing), the states will be verified first -- only neighbor edatas from the same arena and with the same state will be accessed.	2021-03-29 17:19:53 -07:00
Qi Wang	4d8c22f9a5	Store edata->state in rtree leaf and make edata_t 128B aligned. Verified that this doesn't result in any real increase of edata_t bytes allocated.	2021-03-29 17:19:53 -07:00
Qi Wang	70d1541c5b	Track extent is_head state in rtree leaf.	2021-03-29 17:19:53 -07:00
Qi Wang	862219e461	Add quiescence sync before deleting base during arena_destroy.	2021-03-29 17:19:53 -07:00
Qi Wang	61afb6a405	Fix locking on arena_i_destroy_ctl(). The ctl_mtx should be held to protect against concurrent arenas.create.	2021-03-22 23:18:52 -07:00
Qi Wang	3913077146	Mark head state during dss alloc. Specifically, the extent_dalloc_gap relies on the correct head state to coalesce.	2021-03-12 19:17:25 -08:00
Qi Wang	22be724af4	Set is_head in extent_alloc_wrapper w/ retain. When retain is on, when extent_grow_retained failed (e.g. due to split hook failures), we'll try extent_alloc_wrapper as the last resort. Set the is_head bit in that case to be consistent. The allocated extent in that case will be retained properly, but not merged with other extents.	2021-03-12 10:20:08 -08:00
David Goldblatt	73ca4b8ef8	HPA: Use dirtiest-first purging. This seems to be practically beneficial, despite some pathological corner cases.	2021-02-19 15:10:54 -08:00
David Goldblatt	0f6c420f83	HPA: Make purging/hugifying more principled. Before this change, purge/hugify decisions had several sharp edges that could lead to pathological behavior if tuning parameters weren't carefully chosen. It's the first of a series; this introduces basic "make every hugepage with dirty pages purgeable" functionality, and the next commit expands that functionality to have a smarter policy for picking hugepages to purge. Previously, the dehugify logic would never dehugify a hugepage unless it was dirtier than the dehugification threshold. This can lead to situations in which these pages (which themselves could never be purged) would push us above the maximum allowed dirty pages in the shard. This forces immediate purging of any pages deallocated in non-hugified hugepages, which in turn places nonobvious practical limitations on the relationships between various config settings. Instead, we make our preference not to dehugify to purge a soft one rather than a hard one. We'll avoid purging them, but only so long as we can do so by purging non-hugified pages. If we need to purge them to satisfy our dirty page limits, or to hugify other, more worthy candidates, we'll still do so.	2021-02-19 15:10:54 -08:00
David Goldblatt	6bddb92ad6	psset: Rename "bitmap" to "pageslab_bitmap". It tracks pageslabs. Soon, we'll have another bitmap (to track dirty pages) that we want to disambiguate. While we're here, fix an out-of-date comment.	2021-02-19 15:10:54 -08:00
David Goldblatt	154aa5fcc1	Use the flat bitmap for eset and psset bitmaps. This is simpler (note that the eset field comment was actually incorrect!), and slightly faster.	2021-02-19 15:10:54 -08:00
David Goldblatt	271a676dcd	hpdata: early bailout for longest free range. A number of common special cases allow us to stop iterating through an hpdata's bitmap earlier rather than later.	2021-02-19 15:10:54 -08:00
David Goldblatt	d21d5b46b6	Edata: Move sn into its own field. This lets the bins use a fragmentation avoidance policy that matches the HPA's (without affecting the PAC).	2021-02-19 15:10:54 -08:00
David Goldblatt	fb327368db	SEC: Expand option configurability. This change pulls the SEC options into a struct, which simplifies their handling across various modules (e.g. PA needs to forward on SEC options from the malloc_conf string, but it doesn't really need to know their names). While we're here, make some of the fixed constants configurable, and unify naming from the configuration options to the internals.	2021-02-19 15:10:54 -08:00
David Goldblatt	ce9386370a	HPA: Implement batch allocation.	2021-02-19 15:10:54 -08:00
David Goldblatt	cdae6706a6	SEC: Use batch fills. Currently, this doesn't help much, since no PAI implementation supports flushing. This will change in subsequent commits.	2021-02-19 15:10:54 -08:00
David Goldblatt	480f3b11cd	Add a batch allocation interface to the PAI. For now, no real allocator actually implements this interface; this will change in subsequent diffs.	2021-02-19 15:10:54 -08:00
David Goldblatt	bf448d7a5a	SEC: Reduce lock hold times. Only flush a subset of extents during flushing, and drop the lock while doing so.	2021-02-19 15:10:54 -08:00
David Goldblatt	1944ebbe7f	HPA: Implement batch deallocation. This saves O(n) mutex locks/unlocks during SEC flush.	2021-02-19 15:10:54 -08:00
David Goldblatt	f47b4c2cd8	PAI/SEC: Add a dalloc_batch function. This lets the SEC flush all of its items in a single call, rather than flushing everything at once.	2021-02-19 15:10:54 -08:00
Qi Wang	a11be50332	Implement opt.cache_oblivious. Keep config.cache_oblivious for now to remain backward-compatible.	2021-02-11 11:32:01 -08:00
Jordan Rome	8c5e5f50a2	Fix stats for "tcache_max" (was "lg_tcache_max") This opt was changed here: c8209150f9d219a137412b06431c9d52839c7272 and looks like this got missed. Also update the write type to be unsigned.	2021-02-10 23:01:46 -08:00
Qi Wang	041145c272	Report the correct and wrong sizes on sized dealloc bug detection.	2021-02-08 14:42:27 -08:00
Qi Wang	f3b2668b32	Report the offending pointer on sized dealloc bug detection.	2021-02-08 14:42:27 -08:00
David Goldblatt	edbfe6912c	Inline malloc fastpath into operator new. This saves a small but non-negligible amount of CPU in C++ programs.	2021-02-08 14:17:47 -08:00
David Goldblatt	79f81a3732	HPA: Make dirty_mult configurable.	2021-02-04 20:58:31 -08:00
David Goldblatt	32dd153796	HPA: Make dehugification threshold configurable.	2021-02-04 20:58:31 -08:00
David Goldblatt	4790db15ed	HPA: make the hugification threshold configurable.	2021-02-04 20:58:31 -08:00
David Goldblatt	b3df80bc79	Pull HPA options into a containing struct. Currently that just means max_alloc, but we're about to add more. While we're touching these lines anyways, tweak things to be more in line with testing.	2021-02-04 20:58:31 -08:00
David Goldblatt	56e85c0e47	HPA: Use a whole-shard purging heuristic. Previously, we used only hpdata-local information to decide whether to purge.	2021-02-04 20:58:31 -08:00
David Goldblatt	dc886e5608	hpdata: Return the number of pages to be purged. We'll use this in the next commit.	2021-02-04 20:58:31 -08:00
David Goldblatt	9fd9c876bb	psset: keep aggregate stats. This will let us quickly query these stats to make purging decisions quickly.	2021-02-04 20:58:31 -08:00
David Goldblatt	da63f23e68	HPA: Track pending purges/hugifies in the psset. This finishes the refactoring of the HPA/psset interactions the past few commits have been building towards. Rather than the HPA removing and then reinserting hpdatas, it simply begins updates and ends them. These updates can set flags on the hpdata that prevent it from being returned for certain types of requests. For example, it can call hpdata_alloc_allowed_set(hpdata, false) during an update, at which point the given hpdata will no longer be returned for psset_pick_alloc requests. This has various of benefits: - It maintains stats correctness during purges and hugifies. - It allows simpler and more explicit concurrency control for the various special cases (e.g. allocations are disallowed during purge, but not during hugify). - It lets allocations and deallocations avoid disturbing the purging and hugification orderings. If an hpdata "loses its place" in one of the queues just do to an alloc / dalloc, it can result in pathological edge cases where very hot, very full hugepages never get hugified (and cold extents on the same hugepage as hot ones never get purged). The key benefit though is that tracking hpdatas to be purged / hugified in a principled way will let us do delayed purging and hugification. Eventually this will let us move these operations to background threads, but in the short term the benefit is that it will let us have global purging policies (e.g. purge when the entire arena has too many dirty pages, rather than any particular hugepage).	2021-02-04 20:58:31 -08:00
David Goldblatt	0ea3d6307c	CTL, Stats: report HPA empty slab stats.	2021-02-04 20:58:31 -08:00
David Goldblatt	bf64557ed6	Move empty slab tracking to the psset. We're moving towards a world in which purging decisions are less rigidly enforced at a single-hugepage level. In that world, it makes sense to keep around some hpdatas which are not completely purged, in which case we'll need to track them.	2021-02-04 20:58:31 -08:00
David Goldblatt	99fc0717e6	psset: Reconceptualize insertion/removal. Really, this isn't a functional change, just a naming change. We start thinking of pageslabs as being always in the psset. What we used to think of as removal is now thought of as being in the psset, but in the process of being updated (and therefore, unavalable for serving new allocations). This is in preparation of subsequent changes to support deferred purging; allocations will still be in the psset for the purposes of choosing when to purge, but not for purposes of allocation/deallocation.	2021-02-04 20:58:31 -08:00
David Goldblatt	061cabb712	HPA stats: report retained instead of inactive. This more closely maps to the PAC.	2021-02-04 20:58:31 -08:00
David Goldblatt	d3e5ea03c5	HPA: Track dirty stats.	2021-02-04 20:58:31 -08:00
David Goldblatt	68a1666e91	hpdata: Rename "dirty" to "touched". This matches the usage in the rest of the codebase.	2021-02-04 20:58:31 -08:00

1 2 3 4 5 ...

1772 Commits