Models pick words by predicting what usually comes next, and their training data is full of formal, published prose where delve is common. The safe, average pick wins every time, so delve surfaces far more than it does in normal speech.
A person would say look into or dig into. The model reaches for the fancier option out of habit, which is exactly what makes it a tell.
