Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
97 tokens/sec
GPT-4o
53 tokens/sec
Gemini 2.5 Pro Pro
44 tokens/sec
o3 Pro
5 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

cuPC: CUDA-based Parallel PC Algorithm for Causal Structure Learning on GPU (1812.08491v4)

Published 20 Dec 2018 in cs.DC, cs.LG, and q-bio.QM

Abstract: The main goal in many fields in the empirical sciences is to discover causal relationships among a set of variables from observational data. PC algorithm is one of the promising solutions to learn underlying causal structure by performing a number of conditional independence tests. In this paper, we propose a novel GPU-based parallel algorithm, called cuPC, to execute an order-independent version of PC. The proposed solution has two variants, cuPC-E and cuPC-S, which parallelize PC in two different ways for multivariate normal distribution. Experimental results show the scalability of the proposed algorithms with respect to the number of variables, the number of samples, and different graph densities. For instance, in one of the most challenging datasets, the runtime is reduced from more than 11 hours to about 4 seconds. On average, cuPC-E and cuPC-S achieve 500 X and 1300 X speedup, respectively, compared to serial implementation on CPU. The source code of cuPC is available online [1].

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (4)
  1. Behrooz Zarebavani (2 papers)
  2. Foad Jafarinejad (1 paper)
  3. Matin Hashemi (12 papers)
  4. Saber Salehkaleybar (41 papers)
Citations (30)

Summary

We haven't generated a summary for this paper yet.