I'd rather stick with InstructEmbedding instead of pandering to flavour of the month LLM. That way I keep my key components insulated from drastic changes
LLM inputs are the worst candidates for caching. The only place where caching might make sense if you have a public facing service & have coupled it with a vector cache instead of a typical word for word caching of the prompt