Hugging Face
Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up
1
Michael Leonardo Aguas
sensei-ml
Follow
drak-hf's profile picture
1 follower
·
2 following
sensei-ml
michael-leonardo-aguas-37b692206
AI & ML interests
Computer Vision Natural Language Processing Large Language Models
Recent Activity
Reacted to
louisbrulenaudet
's
post
with ❤️
about 2 months ago
My biggest release of the year: a series of 7 specialized embedding models for information retrieval within tax documents, is now available for free on Hugging Face 🤗 These new models aim to offer an open source alternative for in-domain semantic search from large text corpora and will improve RAG systems and context addition for large language models. Trained on more than 43 million tax tokens derived from semi-synthetic and raw-synthetic data, enriched by various methods (in particular MSFT's evol-instruct by @intfloat), and corrected by humans, this project is the fruit of hundreds of hours of work and is the culmination of a global effort to open up legal technologies that has only just begun. A big thank you to Microsoft for Startups for giving me access to state-of-the-art infrastructure to train these models, and to @julien-c, @clem 🤗, @thomwolf and the whole HF team for the inference endpoint API and the generous provision of Meta LLama-3.1-70B. Special thanks also to @tomaarsen for his invaluable advice on training embedding models and Loss functions ❤️ Models are available on my personal HF page, into the Lemone-embed collection: https://huggingface.co/collections/louisbrulenaudet/lemone-embed-66fdc24000df732b395df29b
updated
a Space
about 2 months ago
sensei-ml/crack_detection_for_structural_health_monitoring
liked
a model
about 2 months ago
sensei-ml/concrete_crack_images_classification
View all activity
Organizations
sensei-ml
's activity
All
Models
Datasets
Spaces
Papers
Collections
Community
Posts
Upvotes
Likes
liked
a model
about 2 months ago
sensei-ml/concrete_crack_images_classification
Image Classification
•
Updated
Oct 1
•
7
•
1