Moore's Law doesn't really work anymore. It hasn't held for the last 4 years, IIRC. That's why a lot of focus has been on multicore systems to try and get new speedups based on parallel computing as opposed to getting purely single-chip improvements.
They are, but performance in teraflops is not the same as performance in general computing activities. We are now scaling horizontally within a single cpu package.