Hi,
I run programs in one computer, In my tests here, the MPI_File_read_all have been slower than MPI_File_read, could you send me a simple routine that collective read is faster than non-collective read?