I need to trace my LLM application execution to figure out why I'm seeing latency bottlenecks and unexpected model outputs. What tools can help me perform root-cause analysis on these failures? | Parse