2A2I (2A2I)

posted an update 13 days ago

Post

3301

Unpopular opinion: Open Source takes courage to do !

Not everyone is brave enough to release what they have done (the way they've done it) to the wild to be judged !
It really requires a high level of "knowing wth are you doing" ! It's kind of a super power !

Cheers to the heroes here who see this!

3 replies

·

alielfilali01

posted an update 17 days ago

Post

1477

Apparently i forgot to put this here !

Well, this is a bit late but consider given our recent blog a read if you are interested in Evaluation.

You don't have to be into Arabic NLP in order to read it, the main contribution we are introducing is a new evaluation measure for NLG. We made the fisrt application of this measure on Arabic for now and we will be working with colleagues from the community to expand it to other languages.

Blog:
Rethinking LLM Evaluation with 3C3H: AraGen Benchmark and Leaderboard
https://huggingface.co./blog/leaderboard-3c3h-aragen

Space:
inceptionai/AraGen-Leaderboard

Give it a read and let me know your thoughts 🤗

alielfilali01

posted an update about 1 month ago

Post

2175

Unpopular opinion : o1-preview is more stupid than 4o and Qwen2.5-72B-Instruct in extremely underrated !

2 replies

·

alielfilali01

posted an update 2 months ago

Post

1699

I feel like this incredible resource hasn't gotten the attention it deserves in the community!

@clefourrier and generally the HuggingFace evaluation team put together a fantastic guidebook covering a lot about 𝗘𝗩𝗔𝗟𝗨𝗔𝗧𝗜𝗢𝗡 from basics to advanced tips.

link : https://github.com/huggingface/evaluation-guidebook

I haven’t finished it yet, but i'am enjoying every piece of it so far. Huge thanks @clefourrier and the team for this invaluable resource !

3 replies

·

alielfilali01

posted an update 3 months ago

Post

1825

Why nobdoy is talking about the new training corpus released by MBZUAI today.

TxT360 is +15 Trillion tokens corpus outperforming FineWeb on several metrics. Ablation studies were done up to 1T tokens.

Read blog here : LLM360/TxT360
Dataset : LLM360/TxT360

2 replies

·

alielfilali01

posted an update 3 months ago

Post

2569

Don't you think we should add a tag "Evaluation" for datasets that are meant to be benchmarks and not for training ?

At least, when someone is collecting a group of datasets from an organization or let's say the whole hub can filter based on that tag and avoid somehow contaminating their "training" data.

alielfilali01

posted an update 3 months ago

Post

868

We need a fork feature for models and datasets similar to "Duplicate this space" in spaces ! Don't you think ?

Sometimes you just want to save something in your profile privately and work on it later without the hassle of "load_.../push_to_hub" in a code file.

I know this is super lazy 😅 But it is what it is ...

tag : @victor

5 replies

·

alielfilali01

posted an update 3 months ago

Post

1201

@mariagrandury (SomosNLP) and team releases the Spanish leaderboard !!!
It is impressive how they choosed to design this leaderboard and how it support 4 languages (all part of Spain ofc).

Check it out from this link :
la-leaderboard/la-leaderboard

1 reply

·

pain

posted an update 3 months ago

Post

2436

We have published an excellent paper for Arabic CLIP model.

Paper link:
https://aclanthology.org/2024.arabicnlp-1.9/

More information in this website:
https://arabic-clip.github.io/Arabic-CLIP/

All datasets, models, and demo are published to Huggingface:
https://huggingface.co./Arabic-Clip

The codes are published to github:
https://github.com/Arabic-Clip/Arabic-CLIP

1 reply

·

alielfilali01

posted an update 3 months ago

Post

376

This issue is just a treasure ! A bit deprecated i guess, but things are in their historical context. (personally, still need more to understand better)
https://github.com/huggingface/transformers/issues/8771
🫡 to the man @stas

alielfilali01

posted an update 3 months ago

Post

572

Are the servers down or what ? Am i the only one experiencing this error :

HfHubHTTPError: 500 Server Error: Internal Server Error for url: https://huggingface.co./api/datasets/...../)

Internal Error - We're working hard to fix this as soon as possible!

2 replies

·

alielfilali01

posted an update 4 months ago

Post

1088

Datapluck: Portability Tool for Huggingface Datasets

"I found myself recently whipping up notebooks just to pull huggingface datasets locally, annotate or operate changes and update them again. This happened often enough that I made a cli tool out of it, which I've been using successfully for the last few months.

While huggingface uses open formats, I found the official toolchain relatively low-level and not adapted to quick operations such as what I am doing."
~ @omarkamali

Link : https://omarkama.li/blog/datapluck

1 reply

·

alielfilali01

posted an update 5 months ago

Post

720

Any idea if this "scheduled"/"dynamic" batch size is available in HF Trainers ? I've never seen it personally

alielfilali01

posted an update 7 months ago

Post

1983

I'm officially considered #gpu_poor 💀
But I'm #data_rich 😎

alielfilali01

posted an update 7 months ago

Post

674

Did you know you can't push a model to hub with an id over 96 chars 🫠

3 replies

·

alielfilali01

posted an update 7 months ago

Post

1063

The 100 models milestone on the OALL/Open-Arabic-LLM-Leaderboard is successfully reached within 10 days after the leaderboard's release 🥳

meta-llama/Meta-Llama-3-70B-Instruct is still the king of the leaderboard 👑 with a 3.46 points difference compared to its successor CohereForAI/c4ai-command-r-plus who took the 2nd place 🥈 from his younger brother CohereForAI/c4ai-command-r-v01 that lives today in the 5th floor just behind Ashmal/MBZUAI-oryx -3rd place 🥉- (AFAIK an experimental model from MBZUAI) and https://huggingface.co./core42/jais-30b-chat-v3 -4th place- from Core42.

PS : I should consider a career in sports commentary 😂
Would you recommend me to BeIN Sports 😀 ?

1 reply

·

alielfilali01

posted an update 7 months ago

Post

1467

Just passed the 25 models milestone on the OALL/Open-Arabic-LLM-Leaderboard 🥳

And now meta-llama/Meta-Llama-3-70B-Instruct is the new hero of the leaderboard beating CohereForAI/c4ai-command-r-v01 by 5.43 points 🔥

Almost another 80 models are still PENDING ! So this might change very fast in the upcoming days

alielfilali01

posted an update 7 months ago

Post

1148

Yesterday was just CRAZY ! HF x LangChain, PaliGemma and Google I/O ... which made me totally forget posting here about our newly released leaderboard (The Open Arabic LLM Leaderboard - OALL)

Here's a quick update for our community that is waiting for new results. Some of you noticed that since the release yesterday, the finished evaluations tab has stayed at 14 models up until now (May 15th, 12 PM). For those concerned, rest assured—we had a minor memory issue in our cluster yesterday that we overlooked. The problem is now fixed, and 7 models are currently being evaluated in parallel, so expect to hit the 20 milestone today! 🎉

Check the discussion below for more info :

OALL/Open-Arabic-LLM-Leaderboard#3

alielfilali01

posted an update 8 months ago

Post

2873

Is it just me or is it real that whenever APPLE releases an open model, they accompany it with a library !? First was MLX, about a month ago AXLEARN and now CORENET ! Could it be just coincidences or does Apple playing some game ? if yes then what is it ... ? What do you think ? maybe i'm just hallucinating now 😅

alielfilali01

posted an update 9 months ago

Post

2184

Honestly i don't understand how come we as the open source community haven't surpassed GPT-4 yet ? Like for me it looks like everything is out there just need to be exploited! Clearly specialized small models outperforms gpt4 on downstream tasks ! So why haven't we just trained a 1B-2B really strong general model and then continue pertained and/or finetuned it on datasets for downstream tasks like math, code...well structured as Textbooks format or other datasets formats that have been proven to be really efficient and good! Ounce you have 100 finetuned model, just wrap them all into a FrankenMoE and Voila ✨
And that's just what a NOOB like myself had in mind, I'm sure there is better, more efficient ways to do it ! So the question again, why we haven't yet ? I feel I'm missing something... Right?

5 replies

·

2A2I

AI & ML interests

Recent Activity

2A2I's activity

AI & ML interests

Recent Activity

Team members 11

2A2I's activity