EDDY GIUSEPE CHIRINOS ISIDRO, PhD

EddyGiusepe

AI & ML interests

Speech Processing (ASR, Speaker Identification, Text-to-Speech) NLP (text classification, Sentence similarity, Token classification, etc) Computer vision (Image classification)

Recent Activity

upvoted a collection 5 days ago
๐Ÿ“ FineMath
upvoted a collection 5 days ago
ModernBERT
liked a Space 6 days ago
data-agents/jupyter-agent
View all activity

Organizations

Spaces-explorers's profile picture Hackathon Somos NLP 2023: Los LLMs hablan Espaรฑol's profile picture

EddyGiusepe's activity

reacted to davanstrien's post with ๐Ÿ”ฅโค๏ธ 6 days ago
view post
Post
2472
First dataset for the new Hugging Face Bluesky community organisation: bluesky-community/one-million-bluesky-posts ๐Ÿฆ‹

๐Ÿ“Š 1M public posts from Bluesky's firehose API
๐Ÿ” Includes text, metadata, and language predictions
๐Ÿ”ฌ Perfect to experiment with using ML for Bluesky ๐Ÿค—

Excited to see people build more open tools for a more open social media platform!
reacted to davanstrien's post with ๐Ÿค— 6 days ago
view post
Post
489
Increasingly, LLMs are becoming very useful for helping scale annotation tasks, i.e. labelling and filtering. When combined with the structured generation, this can be a very scalable way of doing some pre-annotation without requiring a large team of human annotators.

However, there are quite a few cases where it still doesn't work well. This is a nice paper looking at the limitations of LLM as an annotator for Low Resource Languages: On Limitations of LLM as Annotator for Low Resource Languages (2411.17637).

Humans will still have an important role in the loop to help improve models for all languages (and domains).
reacted to dvilasuero's post with โค๏ธ๐Ÿ”ฅ 6 days ago
view post
Post
2261
๐ŸŒ Announcing Global-MMLU: an improved MMLU Open dataset with evaluation coverage across 42 languages, built with Argilla and the Hugging Face community.

Global-MMLU is the result of months of work with the goal of advancing Multilingual LLM evaluation. It's been an amazing open science effort with collaborators from Cohere For AI, Mila - Quebec Artificial Intelligence Institute, EPFL, Massachusetts Institute of Technology, AI Singapore, National University of Singapore, KAIST, Instituto Superior Tรฉcnico, Carnegie Mellon University, CONICET, and University of Buenos Aires.

๐Ÿท๏ธ +200 contributors used Argilla MMLU questions where regional, dialect, or cultural knowledge was required to answer correctly. 85% of the questions required Western-centric knowledge!

Thanks to this annotation process, the open dataset contains two subsets:

1. ๐Ÿ—ฝ Culturally Agnostic: no specific regional, cultural knowledge is required.
2. โš–๏ธ Culturally Sensitive: requires dialect, cultural knowledge or geographic knowledge to answer correctly.

Moreover, we provide high quality translations of 25 out of 42 languages, thanks again to the community and professional annotators leveraging Argilla on the Hub.

I hope this will ensure a better understanding of the limitations and challenges for making open AI useful for many languages.

Dataset: CohereForAI/Global-MMLU
reacted to suayptalha's post with ๐Ÿ‘๐Ÿ”ฅ 6 days ago
view post
Post
1556
๐Ÿš€ FastLlama Series is Live!

๐Ÿฆพ Experience faster, lighter, and smarter language models! The new FastLlama makes Meta's LLaMA models work with smaller file sizes, lower system requirements, and higher performance. The model supports 8 languages, including English, German, and Spanish.

๐Ÿค– Built on the LLaMA 3.2-1B-Instruct model, fine-tuned with Hugging Face's SmolTalk and MetaMathQA-50k datasets, and powered by LoRA (Low-Rank Adaptation) for groundbreaking mathematical reasoning.

๐Ÿ’ป Its compact size makes it versatile for a wide range of applications!
๐Ÿ’ฌ Chat with the model:
๐Ÿ”— Chat Link: suayptalha/Chat-with-FastLlama
๐Ÿ”— Model Link: suayptalha/FastLlama-3.2-1B-Instruct
reacted to ginipick's post with โค๏ธ๐Ÿ‘€๐Ÿ”ฅ 6 days ago
view post
Post
4144
๐ŸŒŸ Digital Odyssey: AI Image & Video Generation Platform ๐ŸŽจ
Welcome to our all-in-one AI platform for image and video generation! ๐Ÿš€
โœจ Key Features

๐ŸŽจ High-quality image generation from text
๐ŸŽฅ Video creation from still images
๐ŸŒ Multi-language support with automatic translation
๐Ÿ› ๏ธ Advanced customization options

๐Ÿ’ซ Unique Advantages

โšก Fast and accurate results using FLUX.1-dev and Hyper-SD models
๐Ÿ”’ Robust content safety filtering system
๐ŸŽฏ Intuitive user interface
๐Ÿ› ๏ธ Extended toolkit including image upscaling and logo generation

๐ŸŽฎ How to Use

Enter your image or video description
Adjust settings as needed
Click generate
Save and share your results automatically

๐Ÿ”ง Tech Stack

FluxPipeline
Gradio
PyTorch
OpenCV

link: ginigen/Dokdo

Turn your imagination into reality with AI! โœจ
#AI #ImageGeneration #VideoGeneration #MachineLearning #CreativeTech
  • 7 replies
ยท