We already have pretty powerful local LLM models, which are getting uncensored by community to have 0 refusals.
You need special rigs with a lot of VRAM to run best models, but if you're taking it seriously, it's worth it.
Having no rate or tokens limit is huge.