Check out Agenta (https://github.com/agenta-ai/agenta or agenta.ai), it allows you to debug as in you have a playground where you can see the traces for each call you make and run evals on your workflow
Which tool do you think is better? Currently, it seems that langsmith and llmstudio might be useful.