Scale मा, throughput CPU count भन्दा contention, cache effects, र असीमित queues ले बढी सीमित हुन्छ — एउटा बिन्दु नाघेपछि threads थप्दा कुरा ढिलो हुन्छ। Amdahl's र Universal Scalability Law दुवैले यो अनुमान गर्छन्: coordination खर्च अन्ततः हावी हुन्छ।
Contention
धेरै threads एउटै lock वा cache line का लागि लड्दा, तिनीहरू serialize हुन्छन् र काम गर्नुको सट्टा कुर्दै समय बिताउँछन्। Universal Scalability Law अन्तर्गत, contention र coherency खर्च ले optimal thread count नाघेपछि throughput सक्छ।
