Wednesday, April 15, 2009

// // Leave a Comment

Феймлаб 2009 - Лаборатория за слава | FameLab 2009 in Bulgaria


Благодаря на Любов и на Божидар че ми говориха за това и ме насърчиха да участвам. :)



Финал - София, 11. май, 17 ч., Театър София

...

Well... Thanks to Lyubov Kostova and Bozhidar Stefanov for talking me about.
It was like a joke, but I participated and I'm going to the final in Sofia, 11-th May 2009. :)

http://www.britishcouncil.org/bg/bulgaria-beautiful-science-famelab-rules.htm
Read More

Thursday, April 9, 2009

// // Leave a Comment

"Задачата за моториста" - втори трейлър (2:30 мин)

(The bikers problem - English (subtitles) - coming soon)

Лекциите на проф. Математиков

ЕПИЗОД 1.

Задачата за моториста
Трейлър 2

Филм на Тодор Арнаудов

Жанр: Комедия, моноспектакъл, експериментален, сериал
Синопсис: Чудатият проф. Математиков разказва за себе си и обяснява как се решава Задачата на моториста.





Twenkid Studio 2008
http://twenkid.com
Read More

Sunday, April 5, 2009

// // Leave a Comment

Бисери по Микропроцесорна техника (из сп. "Свещеният сметач", 2003)

Abstract: Comic quotes from lectures for Microprocessor studies. An article from my old e-zine "Sacred Computer", 8/2003.


По времето когато писах тази статия нямаше как да намеря подходящи читатели.
Е, вече има...


Не пропускайте и юнашката народна песен в края - "Не ходи, Микропроцесоре"!



"Свещеният сметач" - БРОЙ 26 (септември 2003)

БРОЙ 26 (септември 2003)
Попаднах на "referati-bg.com" и реших да разгледам за интересни теми в областта на "новите високи информационно-комуникационно интерактивни технологии", че да взема да се пообразовам с някУй най-учен лаф.

Сигурно много нашенски студенти по "Компютърни системи и технологии" и нейните посестрими от факултетите по електроника, за които ВУЗ-ът е първото място, където научават за тригери, броячи, шифратори и т.н., започват да се посвещават в тайните на науката от читанки, подобни на тази по-долу, която на места звучи като написана от ученик от горния курс на детската градина.

Ето бе, казвам си на ум, открих още една причина повечето специалисти по информационни технологии да предпочитат "международните термини" и папагалското преписване на думи между различни азбуки.

Така ги учат, дечинята, още в "детската градина на университета".


Цитираните по-долу пасажи са пълни с "детско-градински" синтактични, семантични, пунктуационни и нам к'ви грешки; с тавтологии и циклични изречения, които в края повтарят началото си.

Не си спомням да съм чел по-смешно написана разработка по изчислителна техника.


Срамота.

Преподавателят потропа с показалката си по бялата дъска.
- Моля, прекратете разговорите, колеги. Лекцията започва!



Ето с какво ще се запознаем в този час: (цитатите са с оригинално сбиване)


"Въвеждане в курса

Цел: Да даде знания на студентите за архитектурата, структурата и принципа на действие на цифровите и микропроцесорните устройства за събиране, обработка и представяне на информация, принципът на действие на подсистемите и възлите на ЦМПУ
(логически елементи,броячи,регистри,микропроцесори,памети,входно-изходни устройства ) ,а също така понятие за програмен модел на МПУ и основите на програмиране на микропроцесорни системи.Получените знания по дисциплината трябва да позволят на студентите да анализират конфигурацията на МП си-ми да съставят алгоритми и разработват елементарни програми за МП си-ми за контрол, управление и обработка на информация.

Съдържание на курса

1.Архитектура на ЦМПУ.Архитектура и принцип на действие на идеализиран микропроцесор.
2.Основи на булевата алгебра.Аритметични и логичски операции с числа,представени в двоична форма.Основни интегрални логически схеми, реализиращи законите на булевата алгебра.
3.Цифрови електронни елементи и устройства.Комбинационни цифрови и логически устройства( шифратори,дешифратори,мултиплексори,демултиплексори,буфери,сумато-ри ).Елементи и възли на цифровите устройства,реализиращи функции с памет (тригери,регистри,броячи).
4.МП и МП си-ми.Архитектура на 8 битови МП. Памети,входно-изходни у-ва.Интерфейс на МП с паметите и входно-изходните у-ва.Входно-изходна организация на обмена на данни на МП си-ми с периферни у-ва.
5.Си-ма от инструкции на МП.Програмен модел на МП.Методи на адресация в МП.
(...)

Курсът ще бъде разгледан без използване на сложни математически и теоретични постановки,като се използва за основа архитектурата на 8 битови МП,за които има достатъчно литература и които са залегнали в основата на разработване на съвременните МП.Курса почива на базата на МП тип СМ600(601)."




"ТЕМА # 2 Архитектура на цифрово изчислително устройство"


"Архитектурата на машината на фон Неман използва програма състояща се от команди,записани в паметта на машината и всяка команда съдържа най-малко две части.Първата част е копа на операцията, а втората-операнд.Копа оказва на процесора,какво да направи ( +,-,*,/ ),а операнда показва с кого да го направи, информацията, която се намира в самия процесор."


"Коп" идва не от "коп-ам", а от КОП, Код на ОПерацията.
Така "копът на операцията" означава "кодът на операцията на операцията".

- Абе може би имат предвид микроопперацията?

Уви, не - тема #2 е твърде далеч от толкова отвлечени знания,
като микропрограмиране.

(...)

"копа оказва [натиск] на процесора" и
"операнда показва с кого да го направи,
информацията, която се намира в самия процесор".

Последното ми звучи сякаш е писано от странен чуждестранен студент с малък опит с българския, който трудно помни граматическите правила, но има удивителна памет за изписването на отделните думи и никога не ги бърка...


"ROM-постоянно запомнящо устройство, което се записват програми,които трябва да се помнят постоянно"

(...)

"RAM-оперативно запомнящо устройство,в което се запомнят програми,които не се използват често (потребителски програми) и междинни резултати от извършването на операции."

:-S


"ТЕМА # 3 МП-основни определения.Етапи на развитие.Архитектура на типов опростен МП"


"Първият МП е бил произведен от фирмата “INTEL” през 1971г. (INTEL 404)"

СМ-400 е семейство микропроцесорни схеми, също правено през 70-те,
но не през '71 и не от Интел (4004), а от Института по микроелектроника на БАН.


"Появява се първият REAL МП (?!...).Представител на тази група е МП “INTEL8088”,на базата на който е създаден първия(!) персонален компютър (ПК).За този МП е създадена и първата DOS(!), която и днес все още се използва."

Този, който е писал горното, вероятно е троен дезинформиращ таен агент на IBM, Intel и Microsoft... Не бива да се чудим, тогава, защо е разпространено
мнението, че MS DOS е единственият DOS - това се преподава в университета!!!


Относно агента... Пък, викам, може и да е четворен, и на Borland, щото "REAL МП" може да означава:

Микропроцесор, който може да обработва данните тип "REAL" в Турбо Паскал



Да, много е вероятно, защото 8088 май че е първият микропроцесор, за който е правен Borland Turbo Pascal... %-)



ТЕМА # 6 Представяне на информацията в ЦИУ [цифровите изчислителни устройства]



"1.Представяне на числата в десетична си-ма.

"Най-разпространената си-ма е десетичната.Тя носи своето название от латинската дума “digitize” (пръст)"

(...)

"Си-ма" ми звучи по-скоро като дума на език от китайската група, отколкото като съкращение на гръцката дума "система".

"Digitize" пък, както е известно, е глагол, производен на "пръст" (правя нещо с пръст), а не съществителното име "пръст" (digit).

(...)

"ТЕМА # 20.Полупроводникови запомнящи устройства /памети/"

"Ако можещото свързване на диода е прекъснато,то по съответната линия данни ще се генерира информация “0”,а ако можещото свързване е свързано -“1”."

"PROM-програмируеми памети са памети,които се предоставят на разработчика на ЦУ така,че той сам да може да си програмира паметта."


"1/с оптическо изтриване-OPROM"

Може би искат да кажат:
1 - С оптическо изтриване (OPROM).


"Записването на информация в тези памети става чрез специална представка,която се свързва с персонален компютър."

Например "над", "из", "въз", които се свързват с "персонален компютър" и се получава "надперсонален компютър", "изперсонален компютър" и пр...


"Програмата, която трябва да бъде записана в паметта, предварително се зарежда персонален компютър, тества се и когато разберем,че тази програма е готова за записване в паметта се прехвърля в персоналния компютър в паметта,която е записва."

:-S


"В началото на развитието на МП първите RAM памети са съдържали клетки позволяващи съхраняване на един бит информация в една клетка 4-8 бита и броя на клетките е бил от 1-4."

"I/входни линии/-броят на тези линии е толкова,колкото бита информация може да се запише в една клетка от дадения чип."

(I / входни_линии) / (-броя_на_тези_линии) == колкото бита информация може да се запише в една клетка от дадения чип. %-)


"1/по АЛ се избира адресната клетка
2/след определено време по чип селекта се избира чипа,от който ще се чете
3/по линията ПЗ се подава сигнал за зареждане на кондензатора
4/по R/W се подава сигнал четене и информацията се извежда по линия О"


1/по = АЛ
2/след = определено време по "чип селекта"
3/по = линията
4/по = R/W

Лекцията, колеги, ще завършим с, може би изглеждащата малко странна за вашата специалност, фолклористика.


"Но има едно общо правило съгласно,което поне първите 256 адреса да се отделят за RAM.Това е свързано с не обходи-моста за повишаване бързодействието на паметите. Производителя дава информация как е разположено адресното поле."


Като прочетох последния блестящ абзац от лекцията, който не знам кой идиот ми написа, колеги, си спомних за юнашкото народно произведение, което ще предложа на вниманието ви в заключение.






Не ходи, Микропроцесоре

Не ходи, Микропроцесоре, да обхождаш моста, че са ти изедници задаса
устроили. На адресното поле са те причакали. Крачетата за захранването сакат да ти резнат. Без чедо майчината платка да оставят.

Не ходи, сине, процесоре, да обхождаш моста. Изедници недни краката искат да ти счупят. При майка си да не можеш да доходиш. Самотен без цокъл да тлееш. Корозията бавно да те мъчи. Сърцето ти без време да угасне.

Не ходи, Микропроцесоре, да обхождаш моста. 256-адреса те дебнат сo нули високи. Корпуса сакат да ти съдерат. Маркировката ти да изтрият. Да те захвърлят безимен в някой ъгъл. При точковите диоди незнаен да изгниеш.

Не послушал Микропроцесорът майка си платка. Да обходи моста на път той заминал. С изедници 256 по пътя се срещнал. И борил се храбро в неравната битка. През моста опитал на пряко да мине. От адресите клети някъде да се скрие.

Но мостът бил твърде къс за процесор. Токът бил твърде буен под моста. Напрежението чедото стопило. В разцвета на младостта му го изгорило. И майката платка самотна остала. Самотна да плаче за свидно си чедо...

Из юнашка народна песен от ботевградския край ;-))

Речник:
сърце - тактов генератор



© Тодор Илиев Арнаудов, 31.8.2003
© Twenkid Studio 2009
Read More

Monday, March 30, 2009

// // Leave a Comment

John Travolta, Staying Alive and Dancing

A string of... coincidences lead me to recall this Bee Gee's song:





Then pointed me out the awesome performance of John Travolta in:


Staying Alive (1983)








I want to go to dance too!... Searching for a dance club. :^)





Read More

Monday, March 23, 2009

// // Leave a Comment

What's wrong with Natural Language Processing? Part 2. Static, Specific, High-level, Not-evolving...

By Todor Arnaudov

Independent Researcher - Twenkid Research
Independent Filmmaker - Twenkid Studio
ASIC Engineer - (as of March 2009)

MS Software Engineering, Plovdiv University, 2008
BS Computer Science, Plovdiv University, 2007
Internship in RIILP, University of Wolverhampton, 2007



What's wrong with Natural Language Processing?
Part 2. Static, Specific, High-level and many more!



In brief: Static, Specific, High-level, "Short-chain of intelligent operations", Lack of evolution of the systems, Not enough experimental approach and research for new paradigms

Part I: http://artificial-mind.blogspot.com/2009/02/whats-wrong-with-natural-language.html


Overall, I'm disappointed by the state-of-the-art of NLP. I think the approaches are too shallow, too obvious, too static... And many more...


[Critics] Who the hell are you to be disappointed from "the state-of-the-art"? Who cares? When did you finish school?!


And I think after taking into account all of the "wrong" parts, it is not surprizing that the progress is slow.

We are walking on almost the same road.

The road does go to somewhere, but I think - not right there where me and other are aiming to. OK, call us yet "dreamers".

[Critics] Crazy mad "scientists"!

If the approach doesn't change, I can't see how the "normal" researchers would go there . The same stuff which is not going to evolution is being done over and over again...

...

In the beginning, let me note that I believe that any research or research direction, including mainstream NLP research is supposed to be a mirror, indication, derivation, reduction, output or whatever of operation of mind. This implies that there is *something right* in the paradigm, the paradigm maps some aspects of researchers' mind and is supposed to solve some problems. Science proves this "scientifically" - OK, no doubt that, what I call "mainstream NLP" does solve many problems...

However, one of the basic mistakes I see, is that the abstractions of these problems are at so high-level and so dispersed, that it is impossible to "ignite" the engine to run on its own.
Say, it is not an engine, but just a tool. A tool for another mind to "push buttons" and see the output, not an engine or an organism to be born and to develop.


*** What mainstream NLP is doing wrongly, in my opinion ***


-- Trying to do a reverse engineering of language, starting from words, text and very abstract and not based on "good physics" linguistic constructs.

[Critics] What the hell is "Good physics?"

"Good physics" is a basis that allows you to build an "engine", "ignite" it and make it running on its own. :) (Arnaudov, 2009)

I believe text is one of the reductions of the operation of mind, which is dynamic, based on multiple inputs simulation of multiple virtual universes at different levels of details/abstraction. Some of them are at very low level (say sets of images, sets of "video", sets of relations between sensory inputs). Words are pointers to very low-level models in mind.

Language, viewed as a bunch of static text, could not contain enough information to rebuild-back intelligence in the low level. Low level of mind is massively reduced and "cleaned", when converting to text, mind is reconstructing the missing part using its internal rich models.

-- The focus of the models is output - models are based too closely on text itself and on structures which are derived from the text itself and the output words.

Again - too obvious and cheap. Text is a reduction of mind in action. Language is not just a flat bunch of words with tags and a boring set of numbers for distribution and frequency.

Mind operates with images, relations and dynamics of virtual universes/systems that it simulates, and then reduces representation of this simulation into text, which actually is a system of pointers to the items and rules, which represent the real structures in mind. The structures in mind are dynamic, they are not 1 billion word corpora.

[Critics:] And how those "dynamic models" supposed to be modeled? You're stupid theorist! You don't define concepts you use! You're doing nothing!

I did define some of them, but years ago and OK, in pieces yet only in Bulgarian.
Anyway this reminds me a "practical" AI professor of mine I spoke with a year ago. He didn't make a difference between McCarthy and Minsky (who cares anyway?) and haven't even heard of Numenta or other advanced research directions, like simulation of neocortex columns...

I am a theorist, because I've been busy doing stuff for a living, besides theorizing. In fact I haven't been really theorizing since my teenage years. I'm tired of being only theorist, so be prepared!


-- Lack of freedom in researcher's imagination and lack of will to test more imaginative, complex, growing and dynamic models than obvious flat static relations between "scientifically linguistically proven" structures.

Researchers are walking on the same old "paved" road. Freedom? Imagination? Quotations rule this world - not freedom of imagination.

You want to do a radically new experiment? Get lost! If your research was not based on your supervizor's, on the best known researcher in the field so far etc. - then you're not a scientist and your research is not scientific.

Young researchers are trained that way. Fine - knowledge, history, respect, methodology, etc...
That's good. However, then they become experienced researchers, they quote their own papers, which are quoting the previous ones, which are accepted to be in the right direction.

Of course, revolutionary research is hard to be done in such conditions.

[Critics:] What revolutionary research are you talking about? You stupid dreamer! Come down and step on Earth! Learn the real NLP! Join the mainstream and you will be forgiven!

Who told you that I don't learn it?Thanks. I may take this option, but let me first try to do it differently.

-- Systems are not general, they are created to solve specific abstract problems, defined in terms of words or other very abstract concepts. That's like dealing with the symptoms, not with the cause of the "desease".

-- Models are not only specific, but static.

Machine learning, Naive bayesian etc. - they seem to be models in development. But what are they actually learning?

Probabilities between some set of "symbols" inside a set.

That's fine, but what is done with those probabilities between symbols later?
What those models want to do later with these probabilities?
Can they want to do, and do anything at all?

This is too flat. Words are pointless without doing something else with them - humans use words to make somebody do, imagine or feel something.

The purpose of "probabilities" in real natural language is to cause something different than words.

-- Lack of will and intentions in models. Lack of effectors. Lack of general feedback loops for self-improvement.

[Critics] Will and intentions? "Desire is irrelevant. They are machines!" And Computational Linguistics is not exactly Artificial Intelligence! Don't mix the fields!

Mind needs will and effectors. Otherwise it is not mind, but a mere number cruncher. And a pure number-cruncher architecture would hardly have capabilities of mind.

[Critics] Oh... Intelligent Agents. Bravo! You reinvented the wheel!

Thanks! You're so sweet!

Here we are - another weak part of mainstream research.

All that dividing of everything, instead of integration.

This division of everything is connected with the tendency of mainstream researchers to solve specific dispersed abstract problems, but not to search a solution for general problems which can solve many specific problems in an elegant way. I suggest you check out Boris Kazachenko's site.

What are the pieces and the mechanisms that can build intelligence up? The general mechanisms and evolved system will be capable to solve all anaphora-resolutions, word-sense-disambiguation, multiword expression recognition and whatever...

Let's search for an engine, not for tools.



-- Lack of continuous development and accumulation of experience. Lack of evolution.

Of course. Models are so much hand-crafted and specific, like tricks. Meet some of my unfortunate disappointments in Computational Creativity:

MEXICA: A Computer Model of Creativity in Writing - "Creativity" Disappointment again

Faults in Turing Test and Lovelace Test. Introduction of Educational Test. (Arnaudov, 2007; suggestion of educational test and analysis of works by Bringsjord, S., Ferrucci, D)


These systems (MEXICA and BRUTUS.1) may seem very good at first sight, but after you look under the hood, you will see how much they are based on word-by-word direction and how weak are they in creative generation of text.

These systems really are not "computationally creative", it is implied by the simplicity of the models.

A nice model is growing on its own by communicating with intelligent environment. You shouldn't be capable to understand it in details after it grow. If you are capable to understand the details and follow them - your model is too simple, it is too "young" or both.

If you use to code everything line by line and direct it... If you can predict everything by hand or in an obvious way... Sorry, but this is - at least - very boring!

[Critics] Theorist!

Thank you! :))


-- The following generation of researchers base their work on the work of the previous ones.

Again... Sure, this is science. It should be like that. Of course the state-of-the-art should be known. And one should use the knowledge, accumulated in the past.

However, I think the efforts spent on this is should be dosed.

Instead of imaging and testing new approaches, most of the time typical NLP researchers do study bibles with models which are proven to lead to very painful and slow progress.

Or the bibles consist of solutions which researchers are supposed to implement.


Or researchers are spending long-long time, building hand-crafted tools and databases, which cannot evolve on their own, later on.

The same path for so many years...


[Critics] Slow progress? Parsing, "Marsing", Syntax 45.4%, 67.4%, POS-Tagging: 96.4%, ...

So...? This progress doesn't lead to intelligent machines.
Those numbers do not map to a genuine general intelligence, but to production of tools.

Hand-crafted tricks with text... If you call this "Natural language processing" - OK, it's great.
This is useful to a certain degree and for particular class of problems.

Yes, mainstream NLP at the moment:

- Is useful.
- Solve some abstract specific problems by heuristics.
- It works to some degree for "intelligent" tasks, because of course language do maps mind.

However, the mainstream still does not lead to a chain of intelligent operations, there are not loops and cumulative development.


-- The lenght of the chain of inter-related intelligent operations in NLP today is very short. This is related to the lack of will and general goals of the systems. These systems are "push-the-button-and-fetch-the-result".

-- Swallowing of a huge corpus of 1 billion of words or so and a computation of statistical dependencies between tokens is not the way mind works.

!!! Mind learns step by step, modeling simpler constructs/situations/dynamics/models before reaching to more complex.
!!! Temporal relations of the input with different complexity is important.
!!! Mind usually uses many sensory inputs while learning. Very important.
!!! Mind has will, uses feedback and can actively and evolutionary test and improve correctness and effectiveness of its operation, including natural-language-related.


I suggest:

1. Hollistic approach - the goal is building an operational mind with long chain of intelligent operations, not completion of a table with values 94.55% 96.5% 90.4% and a long list with quotes in the end of a paper.

2. System must have will and effectors and must evolve. And saying "to evolve", I am not talking about "genetic algorithms", I'm talking about increasing of complexity by fetching "complexity" from the environment and pushing it into the system.

In other words:

Methodology for building very complex systems:

-- Don't do everything by hand, design something which is capable to design parts of it on its own
.

3. If doing reverse engineering - let it be reverse engineering in the beginning of mind development and reverse engineering of the evolution of mind. Not reverse engineering of text.

4. Straight Engineering. Experimental engineering. Experimenting with designs of systems which evolve and fetch complexity.


[Final Critics] Who the hell are you, crazy ignorant stupid kid? "50 years of research of the brightest, talented, etc!" And you think you will change the world! Crazy!

If you are walking on the wrong way, you can't reach the right place, even if you were "the brightests" etc. persons. They couldn't find the right way.

I think one of the important issues with NLP research is that it had been lacking persons with the appropriate combination of talents, mindset and personality to go to a different path than those 50-years old one.

It is not easy to state: "I think this is a wrong approach, let's find another one!", especially if you are young.

Most researchers accept "this is correct, because - quote prof. A, prof. B... They are from University C, which has the most publication in journals D, E and F, which are ... (Oops, there are no Nobel prizes in NLP).

Anyway - therefore, this is the best, because it is quoted there and has 89.95% in this measure, which is accepted by.... Also, the paper suggests 94.34% in the test of "interrelated multipart tagging of coverage structures" etc., so this is real!." and so on.

Or they just want to have their PhD now, and the easiest and fastest way is to fetch a topic from the mainstream and do it the way it is done - these topics are... In Bulgarian it's called "Dissertabilni" - acceptable for a PhD. But mainstream is supposed to be behind the cutting edge.


So I'll say it again:

The reason why so many researchers are doing the same research and progressing so slowly is that they do assume that the others with higher status are right, and base their "original" research too much on it. They do not imagine wildly enough.

The same trivial, not original, not really inventive research, dealing with the old obvious parameters and items, supposed to be "the right ones"....


Conclusion: The paradigm of NLP is wrong.


THE END


To be continued...


Best Regards
Todor Arnaudov


Suggested reading (google): Boris Kazachenko, Jeff Hawkins, Todor Arnaudov (български - http://eim.hit.bg/razum)
Read More

Thursday, March 19, 2009

// // 2 comments

"The Wedding" (short film, comedy) - trailer with subtitles

Trailer of one of my new short films - now with subtitles.
It is in post-production, have to finalize the editing.

Synopsys: A young reporter is going to a mass Wedding ceremony of the Sect of FMI in the University of Plovdiv. Everything is funny, until he realizes that he is also supposed to be marrried and he has no idea who...






Twenkid Studio


If subtitles don't appear - click on the right down on the youtube window, CC.
Read More

Tuesday, March 10, 2009

// // Leave a Comment

Сватбата (комедия) - трейлър | The Wedding (film, comedy, trailer)



Twenkid Studio Film: A trailer of my new short film - "The Wedding". Subtitled!

Synopsys: A young reporter is at a mass Wedding ceremony in the University of Plovdiv. Everything is funny, until he realizes that he is also supposed to be marrried and he has no idea who...

Trailer (subtitled): http://www.youtube.com/watch?v=RaQD5HPVK-0


...
Сюжет: ... Усмихнат репортер предава от една голяма, масова сватба на Сектата на ФМИ, която се провежда във Факултета по математика и информатика на Пловдивския университет... Преди да разбере, че той също е плануван за Сватбата...
Жанр: Комедия

http://twenkid.com/films.htm#svatbata

Сценарий, режисура, оператор, монтаж и в главната роля: Тодор Арнаудов

Участват още:

* Никола Вълчанов
* Иванка Николова
* Димитър Благоев
* Владимир Шкуртов
* Марта Манолова
* Димитър Мекеров
* Запрян и Ивета
* Ивелина, Марина, Ани, Деян, Радослав
* Диляна, Любен, Наско, Иван Минов
* и мн. др.

© Twenkid Studio 2009


Сватбата (комедия) - трейлър



Щракнете върху видеото за да го видите с по-високо качество и върху "HQ":






Очаквайте скоро...



http://twenkid.com
http://twenkid.blogspot.com

http://artificial-mind.blogspot.com
Read More

Thursday, March 5, 2009

// // 1 comment

Silent Party, Pamporovo and Twenkid Research

1. Silent Party

The first Silent Disco in Bulgaria, in Plovdiv, organized by Smirnoff. Headphones, controllable volume, two DJs. Cool! I've always dreamed of a disco with a reasonable sound and without smokers. At least the first goal was achieved there...

http://www.badzhakov.com/blog/?p=1031


2. Pamporovo

I was there for a day. It was enough for finding some beautiful locations for one of my short films in production... :)


3. Twenkid Research


I've just put on-line the first stub-version of Twenkid Research web site. Enjoy.


4. Twenkid Studio's Blog

Started to make it on blogspot: http://twenkid.blogspot.com. Design template not fixed yet...
Read More

Thursday, February 26, 2009

// // Leave a Comment

Christopher Riley and Michael Pitts in NATFIZ, Sofia!



Благодаря на моя приятел Божидар Стефанов и участник във FameLab - Лаборатория за слава, че ми каза; на Любов от Британския съвет и на Малкълм Лав за организацията на това събитие. На НАТФИЗ за залата... И най-вече на лекторитеза вълнуващото изживяване...







Read More

Thursday, February 12, 2009

// // Leave a Comment

Сайтът на Twenkid Studio - Филмовото студио на Тош | Twenkid Studio Web Site



Pictures and elements from the design of Twenkid Studio web site. Created by Todor Arnaudov


"Photographer's Autumn Love" (Autumn Love) - a  picture story by Todor Arnaudov, 2008Поредицата рисунки "Любовта на фотографа" (Есенната любов на фотографа): Ноември, Декември, Януари


I'm working on the design an implementation of the web site of my emerging independent film studio - Twenkid. :)

Check out current version: http://twenkid.com

...

Рисувам си сайта...

Страницата на Twenkid Studio - Филмовото студио на Тош -  http://twenkid.com

Реших да бъде нарисуван. :)

Заглавката ще има още версии, има още жанрове, които засега не са отбелязани. Евентуално ще вмъкна търсене, смятам да направя нещо като imdb:

Twenkid Studio Movie Data Base или Twenkid Movie Data Base- TSMDB или TMDB, където ще се търсят профилите на филми, творци и участници в продукции на студиото.

Има още да се рисува и да се пълни. Засега блогът ще е тук и на http://twenkid.blogspot.com

Ще видим.



Read More

Wednesday, February 4, 2009

// // Leave a Comment

What's wrong with Natural Language Processing?

A short philosophical essay...

NLP?
Do you know about Machine Translation - either rule-based or Statistical, the Lexical Databases like WordNet? Statistical Parsers, POS-Taggers, Parallel Corpora with 1 billion words. Machine Learning, N-grams, Hidden-Markov-Models. And Blah-blah-blah...

What's wrong with NLP?

The main issue I find in the paradigm of NLP today is, I think, embedded in the mindset of the researchers in general, and in the research tradition. IMHO, usually researchers are mathematicians and too nerdy persons.

Usually researchers are mathematicians - not artists, not creative enough and not brave enough to dive into too deep imaginative directions.

Science do also pushes the typical researcher not to invent too much. If he does use his imagination too much and does create "imaginary structures", he might be unable to prove their creation and existence "scientifically" enough to the other members of the sect. He would not be acknowledged etc.

Too nerdy, too mathematical, too obvious and directly "provable" by the raw output data.

As an example of this I would mention Statistical MT.

These distributions work to a certain degree, but this is a trick. It is not really original. There are pure mathematical parameters, which are pretty obvious.

Sorry, but isn't Statistical Machine Translation a mathematical trick?
What are scientific basis of it?

Flat parameters, which can be derived by the distributions of words.

- Take a problem.
- Divide it it into "items" with which you can do something.
- Take the items which you can do something with, and find the combinations which make a difference.
- Take the items and do combinations in order to see what is the difference.

OK, research is a process of exhausting anyway.

Research is exhausting anyway, but if you do not invent structures which are outside and above the obvious ones, you cannot reach too far. However, the relations and combinations which are obvious or easily derived by the raw data without auxiliary "unreal" structures are easier to prove and to be acknowledged in the sect; pardon, I mean science.

This reminds me Quantum Mechanics and the hidden variables. I don't now about physics, but in NLP definitely hidden variables do exist.

Sorry, but in NLP there are hidden variables, out of the text.


I would mention also, that for any creative writer or poet, words are much more than distributions. Writing of a novel is an imaginative process, you are building a world with actors, with laws etc.. Then you simulate that world and record in text what happens.

Creative writing is imagination and simulation, not probability distribution of words.


Actually your dreams about the world you describe are very much more detailed, because the mind is not based on words.

And those Statistical techniques (at least that I know) are flat, because they lack imagination and will, they do not simulate worlds.

Do standard statistical techniques do simulate worlds? Don't think so.

Mind needs to build a virtual world and fit the text to the virtual world, which is simulated during understanding. And I believe that the simulated world in mind is not built by NP, VP etc. NP, VP etc. can be mapped to some aspects of the "simulator" or to modify that simulator in a particular way, but I don't think they are the simulator.

CON: There is no simulator, brain is very slow etc.

(Mostly) memory-based reconstruction is also a simulation. Those 100-level or so of neurons in one "pass" might be pretty enough, if the task is subdivided in a smart way.

Does the "simulator" in mind is built by NP, VP etc.? I don't think so.*

Rather mind can fit its simulator to work with these structures if it needs/wishes.

A prove: babies do not know about words, and in early age when they know something about words, their mind is not based on words anyway. They are pretty intelligent without NP-VP etc., later on they fit NP-VP to something else. I guess so.



To be continued:

- Reflect on explanation of the reason Statistical translation give useful results. ?
- Finally start doing some kind of "simulators" yourself and stop just talking philosophycally! :))


Continues:

What's wrong with Natural Language Processing? Part 2. Static, Specific, High-level, Not-evolving...


* Edit, 1.2.2024: "Do the "simulator"... --> Does the ...; Continue: ... --> Continues

Read More