Tuesday, January 27, 2015

Links of the day 27 - 01 - 2015

Today's links 27/01/2015: #NLP library, Queuing Theory, #NVM in #HPC
Mem0r1es

Monday, January 26, 2015

Links of the day 26 - 01 - 2015

Today's links 26/01/2015: #compression, text extraction, #cloud message queue comparison and #benchmarking
  • ZSTD : short for Z-Standard ( and not Zombie-STD - dirty mind...), is a new lossless compression algorithm, aiming at providing both great compression ratio andspeed for your standard compression needs. It delivers decompression a rate around ~500 MB/s per core and compressiong rate of 200 MB/s ( so no real time for you).
  • Ocropus : when you need to extract text from an image 
  • Comparison of Message Queue Architecture on AWS : the short version is that if you are not latency sensitive and don't want to spend too much effort use SQS for job management and Kinesis for event stream processing. On the contrary if latency is critical ELB or RabbitMQs (or beanstalkds) for job management and Redis for event stream processing.
  • Flotilla : Automated message queue orchestration for scaled-up benchmarking.
Compression, César 1991

Friday, January 23, 2015

Links of the day 23 - 01 - 2015

Today's links 23/01/2015: #cloud storage - Open vStorage, #bitcoin #blockchain scalability, shortest path in corporate communication
  • Open vStorage :  its roadmap and some interesting bits on its architecture.
  • Blockchain scalability : A look at the stumbling blocks to blockchain scalability and some high-level technical solutions.
  • Surprising Facts About Shortest Paths : In a corporate communication network  the shortest path is not the fastest ( Social Network Analysis ). In other words, don’t let your train pass through the central hub for a shortcut, ’cause it’s going to stay there for a long long time.

Thursday, January 22, 2015

Links of the day 22 - 01 - 2015

Today's links 22/01/2015: NUMA & VMs,#openstack & messaging, continuous deliver


Wednesday, January 21, 2015

Links of the day 21 - 01 - 2015

Today's links 21/01/2015: cost of #bigdata scalability and when you are stuck in the middle, to big for small data and to small for big data : #mediumdata
  • Scalability! But at what COST? : Big data systems may scale well, but this can often be just because they introduce a lot of overhead. Rather than making your computation go faster, the systems introduce substantial overheads which can require large compute clusters just to bring under control. In many cases, you’d be better off running the same computation on your laptop. This has been a well know side effect in HPC as for certain type of problem high parallelism just create more overhead than accelerate the processing. The only need for such cluster is the lack of capabilities to buy hardware that is able to fit the necessary amount of data. I guess often people prefer to buy a lot of small cheaper server and say : "look i m running a cluster", than one single beefed up machine that would solve the problem faster.
  • Medium Data : when your big data is too small to warrant a cluster but too big to fit within a single machine .. I think the authors of this article and the one above should collaborate on a follow up one. Also.. one size fit all does not exist.. News at ten.

Tuesday, January 20, 2015

Links of the day 20 - 01 - 2015

Today's links 20/01/2015: Bloom filters, #Unikernel, API framework


Monday, January 19, 2015

Links of the day 19 - 01 - 2015

Today's links 19/01/2015: Data-structure books, #cloud and business agility, Leslie Lamport interview