RL mapping instructions (games)

Κλειστά βάρη Massachusetts Institute of Technology (MIT) 80.9K Παραμέτροι August 2009

Χωρίς εκτίμηση

Δεν υπάρχουν απαιτήσεις υλικού για αυτό το μοντέλο

Τα βάρη αυτού του μοντέλου δεν έχουν δημοσιευθεί, επομένως δεν μπορεί να κατεβεί ή να τρέξει σε δικό σας υλικό οποιουδήποτε μεγέθους. Είναι προσβάσιμο μόνο μέσω του παρόχου του και καμία κάρτα γραφικών δεν το αλλάζει αυτό.

Σε εγγραφή

Πλήρης προδιαγραφή

Όλα τα δεδομένα σχετικά με αυτό το μοντέλο. Τα περισσότερα περιγράφουν πώς εκπαιδεύτηκε παρά πώς τρέχει — χρήσιμο πλαίσιο για την αξιολόγηση του πόσο δουλειά απαιτήθηκε και πώς συγκρίνεται με μοντέλα που έχουν κατασκευαστεί σε διαφορετική κλίμακα.

Προέλευση

Ποιος δημιούργησε αυτό το μοντέλο, πού και πότε δημοσιεύθηκε;

Οργάνωση
Massachusetts Institute of Technology (MIT)
Τύπος οργάνωσης
Academia
Χώρα
United States of America
Δημοσιευμένο
1 August 2009
Συγγραφείς
SRK Branavan, H Chen, LS Zettlemoyer, R Barzilay

Τι κάνει

Οι προβληματικές περιοχές για τις οποίες κατασκευάστηκε το μοντέλο. Ένα μοντέλο μπορεί να περιέχει αρκετές από κάθε μία.

Τομέας
Language
Εντολή
Instruction interpretation

Μέγεθος

Πόσο μεγάλο είναι το μοντέλο και πόσα δεδομένα εκπαιδεύτηκε. Οι παράμετροι είναι το νούμερο που αποφασίζει αν χωράει σε μια δεδομένη κάρτα γραφικών.

Παραμέτροι
80.9K

"We use a policy gradient algorithm to estimate the parameters of a log-linear model for action selection [...] In total, there are 8,094 features [in the Crossblock domain]. [...] This difficulty can be attributed in part to the large branching factor of possible actions at each step — on average, there are [...] 9.78 [actions] in the Crossblock domain"

Δεδομένα εκπαίδευσης
tokens

Shown at beginning of section 7 Total number of documents is 50, average number of actions per document is 5.86 source: https://en.wikipedia.org/wiki/Netflix_Prize

Πώς ταξινομείται

Ετικέτες που εφαρμόζει το πηγαίο σύνολο δεδομένων κατά την παρακολούθηση σημαντικών μοντέλων, και πόσο σίγουρο είναι για την είσοδο.

Αναφορές
318

Πηγές

Από πού προήλθε αυτή η καταγραφή και πότε ελέγχθηκε τελευταία.

Αναφορά
Reinforcement Learning for Mapping Instructions to Actions
Τελευταία ενημέρωση
11 February 2026

Τι σημαίνουν οι αριθμοί

Ιστορικό

RL mapping instructions (games) was published by Massachusetts Institute of Technology (MIT), in United States of America, in August 2009. academia is the category the publisher falls under.

It works in Language, and is recorded as doing instruction interpretation.

Τα βάρη του δεν δημοσιεύθηκαν ποτέ, οπότε μπορεί να αποκτηθεί μόνο μέσω του παρόχου του. Καμία κάρτα γραφικών δεν αλλάζει αυτό.

Απαντήσεις

RL mapping instructions (games) — Συχνές ερωτήσεις

01

What is RL mapping instructions (games) used for?

RL mapping instructions (games) works in Language, and is recorded as handling instruction interpretation. Models frequently carry more than one of each, and the tags describe purpose rather than capability limits.

02

What GPU do I need to run RL mapping instructions (games)?

None. RL mapping instructions (games) is a closed model — its weights were never published, so it cannot be downloaded or run on your own hardware at any price. It is reachable only through its provider.

03

Is RL mapping instructions (games) open source?

The licensing for RL mapping instructions (games) was never recorded in our source data. We treat unstated licensing as closed, because an unrecorded licence is not one to rely on.

04

How many parameters does RL mapping instructions (games) have?

RL mapping instructions (games) has 80.9K parameters. "We use a policy gradient algorithm to estimate the parameters of a log-linear model for action selection [...] In total, there are 8,094 features [in the Crossblock domain]. [...] This difficulty can be attributed in part to the large branching factor of possible actions at each step — on average, there are [...] 9.78 [actions] in the Crossblock domain". That figure is the total, and it is what decides how much memory the model needs — roughly half a gigabyte per billion at the compression most people use.

05

Who created RL mapping instructions (games)?

RL mapping instructions (games) was published by Massachusetts Institute of Technology (MIT), based in United States of America, categorised as academia.

06

When was RL mapping instructions (games) released?

RL mapping instructions (games) was published in August 2009. Capability per parameter has improved considerably since, so a newer model of the same size is often the better use of the same hardware.

Πηγή

Αρχική δημοσίευση

Τελευταία ενημέρωση αρχείου 11 February 2026

Η άλλη κατεύθυνση

Βλέποντάς το από την άλλη πλευρά;

Αυτή η σελίδα ξεκινάει από το μοντέλο. Αν ήδη διαθέτετε μια κάρτα και θέλετε να γνωρίζετε όλα όσα θα τρέξει, Ξεκινήστε από το υλικό αντί..