The RTX 4090 has about the same BF16 Tensor Core TOPs than the A100, assuming 50% MFU (like the A100 40 GB PCIe) it would take 8x longer on 1 RTX 4090 vs 8x A100 80GB SXM, so 12 hours.
Datasheet here for the TOPs https://images.nvidia.com/aem-dam/Solutions/geforce/ada/nvid...
50% MFU should be achievable on the 4090.
I recommend Phind.com, it’s been much better and faster for me than Perplexity Pro. I typically use their custom 70B model but you can also use GPT4 o or Turbo, or Claude 3 Opus.
It takes a while to take down job posts. Everyone likely learned the decision recently. I don’t think the employees who are going to be laid off care about updating the job posts at the moment…