Benhao Tang PRO

benhaotang

AI & ML interests

Physics Master student in theoretical particle physics at Universität Heidelberg, actively looking into the possibilities of integrating AI into future physics research.

Recent Activity

posted an update 1 day ago

Try out my updated implementation of forked OpenDeepResearcher(link below) as an OpenAI compatible endpoint, but with full control, can be deployed completely free with Gemini api or completely locally with ollama, or pay-as-you-go in BYOK format, the AI agents will think dynamically based on the difficulties of given research, compatible with any OpenAI compatible configurable clients(Msty, Chatbox, even vscode AI Toolkit playground). If you don't want to pay OpenAI $200 to use or want to take control of your deep research, check out here: 👉 https://github.com/benhaotang/OpenDeepResearcher-via-searxng **Personal take** Based on my testing against Perplexity's and Gemini's implementation with some Physics domain questions, mine is comparable and very competent at finding even the most rare articles or methods. Also a funny benchmark of mine to test all these searching models, is to trouble shot a WSL2 hanging issue I experienced last year, with prompt: > wsl2 in windows hangs in background with high vmmem cpu usage once in a while, especially after hibernation, no error logs captured in linux, also unable to shutdown in powershell, provide solutions the final solution that took me a day last year to find is to patch the kernel with some steps documented in carlfriedrich's repo and wait Microsoft to solve it(it is buried deep in wsl issues). Out of the three, only my Deep Research agent has found this solution, Perplexity and Gemini just focus on other force restart or memory management methods. I am very impressed with how it has this kind of obscure and scarce trouble shooting ability. **Limitations** Some caveats to be done later: - Multi-turn conversation is not yet supported, so no follow-up questions - System message is only extra writing instructions, don't affect on search - Small local model may have trouble citing source reliably, I am working on a fix to fact check all citation claims

liked a Space 1 day ago

m-ric/open_Deep-Research

liked a model 6 days ago

agentica-org/DeepScaleR-1.5B-Preview

View all activity

Organizations

None yet

benhaotang's activity

upvoted a collection 12 days ago

Paper-to-Read

Collection

6 items • Updated 13 days ago • 3

upvoted an article 13 days ago

Article

Open-R1: Update #1

and 7 others •

15 days ago

• 280

upvoted a paper 14 days ago

Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step

Paper • 2501.13926 • Published 24 days ago • 36

upvoted an article 18 days ago

Article

🐺🐦‍⬛ LLM Comparison/Test: Phi-4, Qwen2 VL 72B Instruct, Aya Expanse 32B in my updated MMLU-Pro CS benchmark

•

Jan 10

• 4

upvoted a collection 25 days ago

DeepSeek-R1

Collection

8 items • Updated 27 days ago • 507

upvoted a collection about 1 month ago

LiveIdeaBench

Collection

3 items • Updated 21 days ago • 5

upvoted a paper about 1 month ago

LiveIdeaBench: Evaluating LLMs' Scientific Creativity and Idea Generation with Minimal Context

Paper • 2412.17596 • Published Dec 23, 2024 • 6

upvoted a collection 3 months ago

v4

Collection

18 items • Updated Oct 20, 2024 • 32