Antes que nada, dejemos en claro que soy un fan de Google. Utilizo la mayoría de sus productos y estoy muy satisfecho con ellos.
El mayor porcentaje de los ingresos de Google viene de la publicidad, lo cual obviamente significa que Google invierte una inmensa cantidad de dinero en hacer que la publicidad sea cada vez más efectiva, básicamente creando un perfil en base a nuestras búsquedas. No tengo nada en contra de eso, pero sí me molesta el orden en que Google presenta los resultados de mis búsquedas. Es decir, si ayudo a un colega con su búsqueda de un teléfono Android; no significa que yo personalmente tenga un interes en esos teléfonos (yo uso Nokia Lumia!). Pero los avanzados algoritmos de Google pueden inferir que yo tengo un interés por ese tópico y darle prioridad en mis futuras búsquedas.
Si te preocupa tu privacidad o si estas cansado de que las búsquedas sean muy "personalizadas", te sugiero que pruebes por una semana este buscador: https://duckduckgo.com. Aquí no se personalizan tus búsquedas, ni tampoco se las guarda.
En este sitio (vale la pena verlo) puedes ver unos ejemplos de cómo Google crea un perfil en base a tus búsquedas.
Comparte tu experiencia.
miércoles, 9 de julio de 2014
lunes, 7 de julio de 2014
From virtualisation to containers, better than electric cars?
Spending a week in Berlin, after my last visit 23 years ago, made me think of the Commodore Amiga 500+. It was the year 91 and I was an exchange student in Germany. The Commodore Amiga was pretty popular there and I was able to test awesome games with my classmates, and off course I ended up buying my favorite ones, like Lemmings. Years passed and my Commodore stopped working, but I still wanted to play some of those great games, but it was not possible to buy an Amiga anymore (it was discontinued in 1992). Then I discovered something new: I could run Amiga games on a PC using something called an Emulator. Although Emulators and virtualisation are not the same, as the guys from Computer World explain here, for me it was the beginning of a journey into emulators and virtualisation.
Some years later a friend of mine had a specialised software that was really hard to configure, and every time his PC or the hard disk crashed (which happened very often) he needed to spend a lot of time and money configuring it all over again. It was then that the curse which haunts all of us who study software engineering (I was still at the university at that time), or anything related to information technology, descended over me: "Hey! you're studying something about computers, solve my problem!" my friend said.
The problem was straight forward: he wanted to configure the operating system and his software for the last time, and then move this "package" (meaning his specialised software and operating system already configured) to a new PC whenever his old one crashed, all done in an easy and practical way. Using the internet I learned about virtual PCs (VPs). To be able to use a virtual PC you need to install a software for the virtualisation, and that's where the magic starts. You execute the virtualisation software and in a window inside your desktop you will see as if a new computer is booting up; in this brand-new computer you need to install a new operating system, software, etc.; exactly as you do with a new physical computer. So, we installed my friend's software in a virtual PC, he could now copy the VP (usually a huge folder) to a new PC every time his old one died. Then he could start the virtual machine that contained his software and would be ready to continue working! Sounds like problem solved, right? well almost; now he was complaining about his software running slower. The solution was to buy more RAM for the PC, because now the hardware was running two operating systems, the base operating system that consumes a lot of memory, and the virtualisation software, which does not require a significant amount of memory by itself, but it contains another operating system called the guest operating system that has same memory requirements as the host.
This is the idea behind virtualisation: multiple virtual computers running on top of one hardware, all sharing and consuming the same physical resources like RAM memory, processor, etc.
It is very practical to have virtual machines that are hardware agnostic since they run on top of any hardware. It is allows to better exploit your hardware by running multiple machines on it; for example, you can have a virtual server for your financial operations, that are heavy during month's end and another virtual server for your logistic operations, that are intense in the middle of the month, this means you will be taking full advantage of your hardware during the whole month.
I recommend reading this article to understand all aspects of virtualisation.
So far I have learned that one of the positive points of virtualisation is that software runs independently of the type of hardware, but on the down side, every virtual machine needs an instance of an operating system that consumes resources. In clouds, where the number of virtual machines is really big, the quantity of resources needed by the guest operating system also become considerable.
Near 2006 Linux introduced a very interesting solution for this problem: containers.
The idea behind containers is: on top of one physical host have only one operating system (no more waste of resources for each operating system on each virtual PC) that can run multiple instances of a program, and do it with certain level of isolation; meaning that each instance of the program believes it is running on a different machine, even with a different network address. Recently an implementation of containers called Docker (http://www.docker.com/) has been in the spotlight because companies like Google and Amazon are contributing to the project, and support this container technology in their own clouds.
The switch from virtualisation to containers, can save the world more energy than switching to electric cars, according to this article published by Wired.
Now that you have a glance of the difference between these 2 technologies, what is your opinion?
domingo, 29 de junio de 2014
Big data: It all started with Google and continues that way...
My 40th birthday is approaching and I start to think what has changed in the last decade specially in IT; this are my reflections about Big Data.
One decade ago, Google's people published some papers detailing a new way to analyze huge stores of information. Data was spread in "small" chunks across thousands of servers. When you asked a question, this query was processed by all those servers in parallel, and you got an answer, usually fast enough. They described this method as MapReduce.
Then Yahoo guys decided to implement MapReduce as an open source project called Apache Hadoop. Now everything related to Big Data is somehow related to Hadoop which has been the hype term for Big Data for some years now. What does Hadoop do? I does MapReduce!
For Hadoop to make sense you have to have some nodes all inter-connected, so when you "ask" something your query is distributed among these nodes.
I think IBM has done a superb job explaining what MapReduce is, you can find it here. It will take 2 minutes to read, and in your next nerd cocktail party you will be able to talk about Hadoop with everyone.
Here is an extract of IBM's explanation:
"As an analogy, you can think of map and reduce tasks as the way a census was conducted in Roman times, where the census bureau would dispatch its people to each city in the empire. Each census taker in each city would be tasked to count the number of people in that city and then return their results to the capital city. There, the results from each city would be reduced to a single count (sum of all cities) to determine the overall population of the empire. This mapping of people to cities, in parallel, and then combining the results (reducing) is much more efficient than sending a single person to count every person in the empire in a serial fashion."
It sounds logical to think that Google processes more information that any other company in the world, and because of that, they create tools to handle huge volumes of data, and those tools seem to be 2-5 years ahead of all others. June 25, 2014 they announced a new technology called Google Cloud Dataflow; in short, Cloud Dataflow is a successor to MapReduce...
Still it is too soon (at least for me) to understand the main differences, but so far I think Big Data started with Google and they are still setting the pace of Big Data.
One decade ago, Google's people published some papers detailing a new way to analyze huge stores of information. Data was spread in "small" chunks across thousands of servers. When you asked a question, this query was processed by all those servers in parallel, and you got an answer, usually fast enough. They described this method as MapReduce.
Then Yahoo guys decided to implement MapReduce as an open source project called Apache Hadoop. Now everything related to Big Data is somehow related to Hadoop which has been the hype term for Big Data for some years now. What does Hadoop do? I does MapReduce!
For Hadoop to make sense you have to have some nodes all inter-connected, so when you "ask" something your query is distributed among these nodes.
I think IBM has done a superb job explaining what MapReduce is, you can find it here. It will take 2 minutes to read, and in your next nerd cocktail party you will be able to talk about Hadoop with everyone.
Here is an extract of IBM's explanation:
"As an analogy, you can think of map and reduce tasks as the way a census was conducted in Roman times, where the census bureau would dispatch its people to each city in the empire. Each census taker in each city would be tasked to count the number of people in that city and then return their results to the capital city. There, the results from each city would be reduced to a single count (sum of all cities) to determine the overall population of the empire. This mapping of people to cities, in parallel, and then combining the results (reducing) is much more efficient than sending a single person to count every person in the empire in a serial fashion."
It sounds logical to think that Google processes more information that any other company in the world, and because of that, they create tools to handle huge volumes of data, and those tools seem to be 2-5 years ahead of all others. June 25, 2014 they announced a new technology called Google Cloud Dataflow; in short, Cloud Dataflow is a successor to MapReduce...
Still it is too soon (at least for me) to understand the main differences, but so far I think Big Data started with Google and they are still setting the pace of Big Data.
sábado, 12 de abril de 2014
¿Por qué Heartbleed es diferente?
Heartbleed no es un virus, ni tampoco un gusano, que son las amenazas de seguridad informática a las cuales estamos acostumbrados. Nos protegemos de ellas simplemente teniendo un antivirus y un sistema operativo actualizado (Windows, por ejemplo). En otras palabras, normalmente lo que tiene un problema de seguridad es nuestro computador. Pero en este caso, es totalmente diferente, son los computadores (servidores) de las grandes compañías los que tienen el problema, compañías como Google, por ejemplo.
No se sabe cuantos y cuales sitios han sido afectados, pero han sido bastantes. Personalmente, creo que más de uno ya es demasiado, aquí hay una lista de los principales sitios afectados.
Ahora entendamos cómo es que tantas compañías fueron afectadas. Un programa informático que es gratuito y de código abierto (creado y mantenido por un grupo de personas independientes) llamado OpenSSL es la causa.
Dado que era gratuito y además considerado de lo más seguro, miles o hasta millones de empresas decidieron utilizarlo para cifrar las conexiones entre nuestros computadores y los de ellos (según la RAE, cifrar se define como: Transcribir en guarismos, letras o símbolos, de acuerdo con una clave, un mensaje cuyo contenido se quiere ocultar). Cada vez que en nuestro navegador web poníamos www.gmail.com y luego escribíamos nuestra contraseña, el programa OpenSSL, en los servidores de Gmail, se encargaba de cifrar el diálogo que existía entre nuestro computador y el de Google, para que nadie más (otros computadores o personas conectadas al Internet, por ejemplo) puedan ver cuál es nuestra clave o el contenido de nuestros correos electrónicos. Pero como suele suceder con programas informáticos, OpenSSL tenía un error, este error recientemente descubierto y bautizado Heartbleed permitía a otras personas ver el contenido de la conversación supuestamente cifrada entre nuestro computador y el de Gmail por ejemplo, esta conversación contiene nuestra contraseña y nuestros emails, en este caso.
Como puedes ver, nosotros como individuos no podemos hacer nada en nuestros computadores personales para subsanar esta vulnerabilidad, pero lo que si podemos hacer es lo siguiente:
- Una vez el proveedor de servicios (Banco, Email, Tienda, etc.) confirme que han configurado en sus servidores la última versión de OpenSSL, la cual no contiene el error, cambiar nuestras contraseñas lo antes posible.
Algunas empresas son muy abiertas y honestas confesando que tienen el problema e informan a todos sus clientes o suscriptores por email acerca de los riesgos, y también confirman cuando ya han subsanado el problema, y que uno debe cambiar la contraseña; no ignoremos estos mensajes y cambiemos nuestras contraseñas.
- También podemos analizar si un sitio tiene el problema o no usando esta herramienta, y si lo tiene, contactarlo y exigir que lo arreglen o cerrar nuestra cuenta.
Como ven, este incidente de seguridad es único y creo que cambiará la historia de la seguridad informática.
Feliz fin de semana cambiando sus contraseñas!
jueves, 9 de enero de 2014
Google's searching capabilities! Looking for a new job for example.
Don’t forget about google’s searching capabilities! Looking for a new job for example?
Google has a lot of powerful searching operators. For example, if your are looking for a new job, and you have noticed that that a lot of companies post those opportunities in taleo.net you can run this query in google: site:taleo.net BW, this will find all BW jobs in taleo.net. Spend some time improving your google search skills for free in this google made course: http://www.google.com/insidesearch/landing/powersearching.html
domingo, 18 de agosto de 2013
La privacidad digital u online
La privacidad digital u online
Las revelaciones, primero de Wikileaks, y
ahora de Snowden (el contratista de la agencia de seguridad de EEUU que reveló
algunos secretos), generaron una serie de preguntas de mis amigos y familiares
sobre el tema de la privacidad digital. Yo no soy ningún experto, pero como
trabajo con informática, es mi “obligación” conocer sobre el tema; esto es algo
común para nosotros los informáticos; nunca falta el amigo o familiar que tiene
un problema con su computador y es “obligación” nuestra tener la respuesta.
Bueno aquí les dejo algunos puntos para
pensar:
Seguridad
básica
Usa contraseñas complejas y que no estén
guardadas en tu computador.
Instala sistemas de protección (antivirus,
etc.).
Siempre actualiza el software de su computador.
(Sobre todo la versión de Java, lo puedes hacer aquí.)
Si tienes un documento con información
delicada guárdalo en un flash drive, no en Dropbox o Google Drive.
Aquí permíteme hacer un paréntesis, yo
crecí y viví hasta mis 30 (sigo en mis 30) en una ciudad tropical muy húmeda,
guardé mis documentos e imágenes importantes en CDs, y luego de 3 años estaban
perdidos por efectos de la corrosión y/o moho. Entonces descubrí los DVDs
con Oro (13 USD por 5 unidades) que prometen una vida de 100 años, no estoy
seguro si es así, pero sé que más de 5 años en una ciudad húmeda sí resistirán.
Entonces, si el documento que quieres conservar no va a cambiar, lo puedes
guardar en uno de estos DVDs en vez de en un Flash drive.
Cifrado
La RAE define cifrar
como:
Transcribir en guarismos,
letras o símbolos, de acuerdo con una clave, un mensaje cuyo contenido se
quiere ocultar.
Con eso ya
tienes la idea de qué significa cifrar, por ejemplo cifrar nuestros correos
electrónicos. Dependiendo de tu grado de interés o si estás interesado en tu
correo personal o el de tu empresa, puedes contratar el servicio de empresas
como www.silentcircle.com o www.voltage.com; si tu presupuesto es 0,
puedes acudir a www.gnupg.org. En esta
última opción tengo que advertirte que vas a invertir un poco de tiempo extra (2
– 4 horas) y además que en tus dispositivos móviles (teléfono móvil, tableta,
etc.) no vas a poder leer tus mensajes cifrados.
Distorsión
Algunos
expertos opinan que usar cifrado, es como enviar un mensaje a las autoridades
diciendo: “Estoy haciendo algo indebido, por eso cifro mis mensajes”,
personalmente no estoy de acuerdo con eso, pero si tu lo estas, una alternativa
al cifrado es la distorsión. www.spammimic.com
por ejemplo, disfraza tu mensaje en lo que parece ser un típico mensaje de Spam.
También puedes combinar esto con www.10minutemail.com,
tus mensajes viven 10 minutos y luego desaparecen para siempre.
Estenografía
No estoy
seguro que este sea el término correcto en español, pues la RAE define
estenografía como taquigrafía, pero no encontré otro mejor para la palabra en
inglés Steganography.
Esta técnica
es para los que están dispuestos a gastar un poco más de tiempo. Básicamente se
trata de enviar mensajes ocultos en fotos, canciones mp3, etc. Por ejemplo, con
secretbook
puedes esconder mensajes en fotos publicadas en el Facebook.
Redes de anonimato
Tor es una red de voluntarios que trata
de mantener anónimas tus actividades como ser que sitios visitas y que
descargas. Es una buena ayuda, pero no estas del todo protegido.
Espero sus
comentarios y preguntas.
Suscribirse a:
Entradas (Atom)
